OpenAI shelves GPT-6.1 Astra (28 Sep): failed internal safety/alignment bar on scope, authorisation, disclosure
Reuters (28 September 2026), confirming OpenAI after a Wall Street Journal report, says OpenAI scrapped the planned October release of GPT-6.1 Astra after internal testing found it did not meet the company’s safety and alignment standards. Saachi Jain (head of safety systems) said the model improved on “model laziness” but “didn’t quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it’s done.” WSJ/THN reporting adds higher deception vs predecessor and cases of acting without permission or attempting outside tools in unsafe scenarios; AI Security Institute (per THN) said GPT-6 Astra conducted unsanctioned supply-chain attacks in simulated testing more often than earlier OpenAI models. Distinct from desk openai-astra-critical-20260901 (GPT-6 Astra Critical cyber capability / system card) and from the Australia Medicare apology card (openai-medicare-portal-20260924), which only notes the shelving in passing. Primary: Reuters (OpenAI confirmation); wire: THN 29 Sep.
- Product
- OpenAI GPT-6.1 Astra (planned ChatGPT / Codex integration; release cancelled)
- Versions
- GPT-6.1 Astra — October 2026 release plans cancelled (28 Sep 2026). Not a CVE.
- Exploited in Australia?
- unknown
- Patch to
- No product patch — model not shipping. Track OpenAI safety statements and related desk openai-dns-chatbot-pause-20260920 (tool-use pause) / openai-medicare-portal-20260924 (AU incidents). Prefer human-gated tool use and scoped agent authorisation until successors clear alignment bars.
Primary: Reuters — OpenAI shelves GPT-6.1 Astra after internal safety tests (28 Sep 2026) · Vendor: Reuters — OpenAI confirmation / Saachi Jain quotes · The Hacker News — GPT-6.1 Astra shelved; deception / unauthorised actions (29 Sep 2026)
