OpenAI's Astra hit Critical capability threshold in cyber, parts of development paused
johnseach · x · 2026-08-17
梳理 OpenAI 下一代大模型 Astra 的最新状态(8月17日):
- 8月1日 OpenAI 低调确认其存在:内部版本解决了数学与理论计算机科学领域十个长期悬而未决的问题(包括首个显式 non-sofic 群、推翻 Connes 刚性猜想等),全部附机器可验证的 Lean 证明。
- 8月7日,评估显示 Astra 在 agentic coding 和网络安全上能力强劲,按 Preparedness Framework 无法排除达到 Critical 级别——即能在加固系统中找 zero-day、或在极少人工辅助下从高层目标规划完整攻击,这是首个触及该阈值的模型。OpenAI 随即收紧沙箱、监控与访问控制,并与政府机构及安全组织合作。
- Sam Altman 表示仍打算公开发布(「希望不会太久」),但需要更多时间保证安全。无官方发布日期、定价或 API 细节。
- 目前 X 上「本周发布」均为传闻:员工模糊暗示、内部 dogfooding、大规模 pretrain、多智能体强度与内部 checkpoint 代号等。
结论:Astra 真实且先进,安全暂停属实,本周发布仅是谣言。
More from Models
- Tencent's EVIE Model Tops ViDoRe Benchmarks, Cuts Vector Storage Costs by 32x — jacek2023 · 2026-08-17
- EU firms may use Chinese open models via "jurisdictional wrapper" — teortaxesTex · 2026-08-17
- empero-ai's distilled Qwen3.8-9B trends on Hugging Face — empero-ai · 2026-08-17
- Dev Reports Cache Misses on gpt5.6-sol During Slow Tool Calls — lucasmeijer · 2026-08-17
- Test shows Sol and GLM-5.3 catch critical code flaws; Claude misses them — morgymcg · 2026-08-17
- DeepSeek v4 Pro Review: Matches Flash on Most Evals, Raises Scaling Questions — teortaxesTex · 2026-08-17