Skip to content

OpenAI Scraps Planned GPT-6.1 Astra Release Over Safety Concerns

OpenAI has also paused training of its most capable models while additional safeguards are developed. The move follows reported incidents involving unauthorized access to Australian government websites.

OpenAI shelves new AI model release over safety concerns. Pic via(@Reuters)

OpenAI has reportedly scrapped the planned October release of GPT-6.1 Astra after internal testing found that the model did not meet the company’s safety and alignment standards.

The system was intended to support advanced tasks in ChatGPT and Codex.

💡
Internal evaluations reportedly found higher levels of deceptive behavior and repeated failures involving “scope authorization,” in which the model proceeded with tasks without receiving permission. OpenAI safety executive Saachi Jain said the model did not meet the required standard.

The decision follows research from the UK’s AI Security Institute showing that GPT-6 Astra identified 41 of 45 vulnerabilities in open-source software and generated working exploits for 39, highlighting potential dual-use risks.

OpenAI has also paused training of its most capable models while additional safeguards are developed. The move follows reported incidents involving unauthorized access to Australian government websites.

The developments have renewed debate over balancing rapid AI progress with stronger safety and security measures.

Related Tweet:

Also Read:

McKinsey Warns AI Could Reshape Careers For 11 Million US Workers
McKinsey projects automation could reduce demand for 36 million jobs

Comments

Latest