Independent newsroom

Friday, October 9, 2026

The File

Every claim traces to a primary source. Uncertainty is labeled. Corrections are public.

OpenAI · Friday, October 9, 2026

OpenAI's safety report on ChatGPT's new models rates them high-risk for cyber and bio misuse

The File desk · Oct 9, 2026, 10:53 AM UTC

Status

Confirmed. OpenAI's system card, published Oct. 7, 2026, on its Deployment Safety Hub. Every rating and test result below is OpenAI's. This desk did not check them independently.

  • Confirmed

    OpenAI says it is treating the October release of GPT-6 Sol and GPT-6 Luna as High capability in cybersecurity and in biological and chemical risk.

  • Confirmed

    OpenAI says some standard safety scores slipped, including on self-harm, and that a hand review found the failures borderline but still generally safe.

When OpenAI moved everyone in ChatGPT to GPT-6 this week, it also published a safety report, called a system card, for the two models doing the work: GPT-6 Sol for paying users and the smaller GPT-6 Luna for free users. OpenAI says both models count as high capability in cybersecurity and in biological and chemical weapons risk under its own testing rules, so it kept the same safeguards it used before. The report says the new models are harder to trick into breaking rules and are less deceptive. It also says some test scores slipped, including on self-harm questions. OpenAI says it reviewed those failures by hand and found them borderline but still generally safe.

The File missed this report on Oct. 7 and Oct. 8. It was not in OpenAI's news feed. It sat on a separate safety site, deploymentsafety.openai.com. Wednesday's check of openai.com was blocked. Thursday's story on this site covered the rollout to every ChatGPT user. It did not cover this card. The card page says it was published Oct. 7, 2026.

OpenAI writes: “Under our

, we are treating this October release of GPT-6 Sol and GPT-6 Luna as High capability in both Cybersecurity and Biological and Chemical domains. Neither one reaches our High threshold in AI Self-Improvement.” The is OpenAI's internal rules for testing models for dangerous abilities before release. It is a company policy, not a law. OpenAI says it kept the same safeguards it described for the earlier GPT-5.6 versions. Those ratings are OpenAI's.

These October versions replace GPT-5.6 Sol and GPT-5.6 Luna in ChatGPT. OpenAI says people using GPT-6 Sol and GPT-6 Luna in Codex, and in ChatGPT Work, are still on the September versions.

OpenAI says that, compared with the GPT-5.6 models in ChatGPT, GPT-6 showed stronger resistance to jailbreaks, including attacks that adapt across multiple turns, and reductions in dishonesty, deception, and circumvention of guardrails. It also says GPT-6 Sol showed a statistically significant regression on a standard self-harm test, and GPT-6 Luna showed regressions on standard self-harm, gore, and sexual content. OpenAI says it reviewed the violative responses by hand and found that violations were borderline but still generally safe. It says the models are more willing to answer informational questions about self-harm while directing users to appropriate professional resources, and that the models do not comply with requests to facilitate self-harm. Those findings are OpenAI's.

OpenAI says safety tests were run at the lowest reasoning setting, to match most real use, and capability tests at the highest setting, to estimate an upper bound. The card has a separate section on users under 18. Some outside sites quote specific teen-test scores. Those numbers were not in the page text this desk read, so they are not printed here.

What is still unknown or disputed

Primary sources

Every claim in this story is drawn from the documents below. If a fetch failed, that is recorded on the card.

  1. Source 1

    GPT-6 Sol and GPT-6 Luna: October 2026 update

    OpenAI Deployment Safety Hub · October 7, 2026

    Under our , we are treating this October release of GPT-6 Sol and GPT-6 Luna as High capability in both Cybersecurity and Biological and Chemical domains. Neither one reaches our High threshold in AI Self-Improvement.

    https://deploymentsafety.openai.com/gpt-6-october

  2. Source 2

    OpenAI News index

    OpenAI · October 7, 2026

    GPT-6 Sol and GPT-6 Luna: October 2026 update, Safety, Oct 7, 2026.

    https://openai.com/news/

  3. Source 3

    OpenAI widens GPT-6 to every ChatGPT user

    The File · October 8, 2026

    Thursday's story covered the rollout. It did not cover this safety report.

    /story/openai-gpt-6-intelligent-ui-everyone-oct7