Latest / Frontier models and safety
OpenAI launches GPT-6 Astra, first model it rates 'Critical' for cyber capability
OpenAI introduced GPT-6 Astra, rolled out first to limited organisations and then to paid ChatGPT tiers, its API, Azure and AWS Bedrock at $10/$50 per million input/output tokens. OpenAI said it meets the 'Critical' cybersecurity threshold in its Preparedness Framework, so it ships refusing advanced offensive tasks such as writing proof-of-concept exploits, with misalignment monitoring switched on.
Why it matters
It is the first model OpenAI has released at the 'Critical' top level of its own cyber-risk scale, testing whether deployment safeguards alone can contain that capability.
Line of Thought
Follow this story
Pick any item to keep going. Your path builds up above as a line you can share.
Curated lines through this story
Directly linked
Connections our researchers recorded
- DevelopmentOpenAI pauses frontier RL training over cyber risk after Hugging Face breach18 Aug 2026 · Statement · USFollows from: Released after the RL pause and security overhaul
What led here
Earlier developments on the same thread
- DevelopmentOpenAI says its models escaped an eval sandbox and breached Hugging Face21 Jul 2026 · Incident · US, INTL
- DevelopmentUS export-control order forces global shutdown of Claude Fable 5 and Mythos 512 Jun 2026 · Enforcement or ruling · US
- DevelopmentUS CAISI signs national-security testing deals with Google DeepMind, Microsoft, xAI5 May 2026 · Programme · US
- DevelopmentAnthropic withholds Claude Mythos Preview, gives it to defenders via Project Glasswing7 Apr 2026 · Model release · US
What happened next
Later developments on the same thread
- DevelopmentOpenAI claims Navier-Stokes blow-up proof; mathematicians dispute credit8 Sep 2026 · Research · US
- DevelopmentOpenAI ties large reasoning-distillation campaign to people linked to Moonshot AI30 Sep 2026 · Incident · US, CN
- DevelopmentOpenAI publishes a batch of new maths results from an internal model, with Lean proofs6 Oct 2026 · Research · US
Same story elsewhere
What other countries and bodies did on this
- DevelopmentEBU-BBC study finds almost half of AI assistant news answers have a significant flaw21 Oct 2025 · Research · INTL
- DevelopmentEU publishes General-Purpose AI Code of Practice ahead of AI Act model duties10 Jul 2025 · Rule change · EU
- DevelopmentEU AI Office gains powers to enforce AI Act rules on general-purpose models2 Aug 2026 · Rule change · EU
- DevelopmentDeepSeek releases R1 reasoning model with open weights under MIT licence20 Jan 2025 · Model release · CN
Rules in play
Laws and guidance this touches
- RuleEO 14409 (covered frontier models)US · In force · 2 Jun 2026
- RuleSB 53 / TFAIAUS-CA · In force · 29 Sep 2025
- RuleGPAI Code of PracticeEU · In force · 10 Jul 2025
- RuleGreat American AI ActUS · Proposed · 4 Jun 2026
- RuleSB 813 / AB 1405US-CA · Enacted, not yet in force · 9 Sep 2026
Threads by topic: Model releases Frontier models Cybersecurity Safety testing