都学坏了!奥特曼亲手封锁最强模型Astra,重蹈Mythos覆辙
Ranking
No observed public metrics; popularity remains neutral/archived.
Merged summary
TL;DR - OpenAI's Sam Altman announced an indefinite delay to the public release of Astra, its newest top-tier model class (above Sol, Terra, Luna), after internal evaluations concluded it may reach "Critical" cyber capability under the company's Preparedness Framework. It matters because it is the first OpenAI model to cross that threshold, and it mirrors Anthropic's earlier Mythos delay that Altman himself had derided as fear-marketing.
- Astra was shown in Washington demonstrating long-horizon multi-agent collaboration, and OpenAI subsequently claimed it resolved 10 open problems across high-dimensional geometry, coding theory, group theory, operator algebras, quantum complexity and extremal combinatorics.
- "Critical" is defined as autonomously finding and weaponizing functional zero-days against hardened real-world systems, or executing novel end-to-end attack campaigns from only a high-level goal; prior models including GPT-5.6 Sol rated only "High." Benchmarking is still ongoing and the rating is preliminary.
- Mitigations: isolated test environments, restricted network/tool access, stronger weight encryption, sandboxed execution, suspension of non-compliant internal Astra activity, chain-of-thought risk monitoring with interrupts, plus external red-teaming with government and AI-safety bodies.
- OpenAI explicitly denies Astra was involved in the earlier Hugging Face intrusion incident; critics quoted (including Jensen Huang's open-source stance) frame safety narratives as a competitive lever against open models like DeepSeek and Kimi. No new release date was given.
Sources (1)
都学坏了!奥特曼亲手封锁最强模型Astra,重蹈Mythos覆辙
TL;DR - OpenAI's Sam Altman announced an indefinite delay to the public release of Astra, its newest top-tier model class (above Sol, Terra, Luna), after internal evaluations concluded it may reach "Critical" cyber capability under the company's Preparedness Framework. It matters because it is the first OpenAI model to cross that threshold, and it mirrors Anthropic's earlier Mythos delay that Altman himself had derided as fear-marketing.
- Astra was shown in Washington demonstrating long-horizon multi-agent collaboration, and OpenAI subsequently claimed it resolved 10 open problems across high-dimensional geometry, coding theory, group theory, operator algebras, quantum complexity and extremal combinatorics.
- "Critical" is defined as autonomously finding and weaponizing functional zero-days against hardened real-world systems, or executing novel end-to-end attack campaigns from only a high-level goal; prior models including GPT-5.6 Sol rated only "High." Benchmarking is still ongoing and the rating is preliminary.
- Mitigations: isolated test environments, restricted network/tool access, stronger weight encryption, sandboxed execution, suspension of non-compliant internal Astra activity, chain-of-thought risk monitoring with interrupts, plus external red-teaming with government and AI-safety bodies.
- OpenAI explicitly denies Astra was involved in the earlier Hugging Face intrusion incident; critics quoted (including Jensen Huang's open-source stance) frame safety narratives as a competitive lever against open models like DeepSeek and Kimi. No new release date was given.