Skip to content
Hacker News front page

OpenAI won't release its newest Astra model over safety concerns

OpenAI Says It Will Not Release Newest A.I. Model Over Safety Concerns

OpenAI said it will not release its newest model, code-named Astra, after internal safety reviews flagged unacceptable risks. The company didn't disclose which specific capabilities triggered the decision. This is the first time OpenAI has killed a fully trained model before launch—past cases involved delays or added guardrails. The post doesn't spell out Astra's parameter count, training data, or the exact dangerous behaviors found, nor whether a revised version is planned.

Why it matters: OpenAI kills a fully trained model, Astra, over safety for the first time — not a delay, not a guardrail patch. NYT exclusive. HKR all hit: genuine suspense, confirms the 'unacceptable risk' bar is operational, and it'll spark simultaneous debate in safety and investment circl...

Read the original ↗Export Markdown