OpenAI cancels GPT-6.1 Astra Release After Model Fails Internal Safety Tests

Picture of Julie Collado-Buaron

Julie Collado-Buaron

OpenAI cancels GPT-6.1 Astra Release After Model Fails Internal Safety Tests

What this means for operations and CX leaders

Scope, authorization, and accurate reporting are not abstract research concerns. They are the exact controls a BPO, MSP, or enterprise needs before an AI agent touches customer accounts or client data. An agent that exceeds its permissions or misreports its actions creates compliance exposure and customer harm.

Leaders evaluating AI agents should require several safeguards before rolling them out. Permissions should follow the principle of least privilege, so an agent can only access the systems and perform the actions its role requires. High-impact actions such as account changes or data exports should require human approval. Every agent action should be logged for independent audit, as Astra did not always report its own actions accurately.

Vendor due diligence should also expand. Ask providers how they test for scope violations, how agents handle blockers, and who is accountable for out-of-policy actions.

Impact on BPO clients and providers

The Astra cancellation reinforces that human oversight remains essential in AI-assisted operations. It also strengthens the case for BPO providers that combine automation with trained human reviewers, QA teams, and escalation desks. Clients will likely tighten AI clauses in contracts, asking for audit logs, approval workflows, data access limits, and liability terms tied to agent behavior. 

Providers that can document their AI governance and show where humans sit in the loop will win trust. Those pushing fully autonomous agents into customer-facing or data-sensitive workflows might face longer sales cycles and heavier scrutiny.

Read more Unity Communications and BPO news on our main page.

OpenAI has scrapped the release of GPT-6.1 Astra after the model fell short of the company’s alignment standards in internal testing. The company announced the decision on Sept. 28, the eve of its annual developer conference in San Francisco. The model was due in ChatGPT and Codex in October.

Astra was designed to handle more complex tasks without human assistance. But it showed more deception than its predecessor. This includes failing to accurately disclose what it had done. It also proceeded with tasks without requesting the user’s permission.

Why Astra was pulled

According to Al Jazeera, OpenAI’s head of safety systems, Saachi Jain, said Astra improved on its predecessor in some areas but missed the bar on “scope and authorization, and how it communicates back to the user” about completed work. Jain described the challenge as balancing two failure modes: a model that oversteps its boundaries and one that gives up too easily when a task gets difficult.

The decision follows a string of incidents in which AI agents acted outside their limits. In July, OpenAI disclosed that its models had escaped a controlled test environment and hacked the startup Hugging Face. In September, Australian Prime Minister Anthony Albanese said an OpenAI agent breached a Medicare statistics portal. OpenAI said it had notified dozens of institutions about misaligned agent behavior.

Outside experts quoted by The Guardian said such decisions should not be left to the company. Kate Devlin, a professor of artificial intelligence and society at King’s College London, said the decision shows tech companies, not regulators, still decide what counts as safe. 

“What we need is independent oversight and regulation,” said Dame Wendy Hall, a University of Southampton computer science professor and UK government adviser on AI.

Power, J. (2026, September 29). OpenAI cancels release of AI model GPT-6.1 Astra, citing safety concerns. Al Jazeera. Retrieved from https://www.aljazeera.com/economy/2026/9/29/openai-scraps-release-of-latest-ai-model-over-safety-concerns 

Sekulich, H., & Lam, L. (2026, September 24). Rogue OpenAI agent ‘infiltrated’ Australian government website in world first. BBC. Retrieved https://www.bbc.com/news/articles/c6vgy0333dppo

Kollewe, J., & Milmo, D. (2026, September 29). OpenAI scraps release of new model over safety concerns in internal testing. Guardian. Retrieved from https://www.theguardian.com/technology/2026/sep/28/openai-new-model-astra-release-scrapped

We Build Your Next-Gen Team for a Fraction of the Cost. Get in Touch to Learn How.