Skip to content

AI Security Concerns Grow As OpenAI Restricts New Model Testing

The move comes after several AI security incidents.

Photo by Levart_Photographer / Unsplash

OpenAI has tightened controls around its unreleased Astra model after preliminary testing showed it could potentially reach a critical level of cyber capability.

The company said it could not rule out Astra being capable of launching sophisticated cyberattacks autonomously. OpenAI has halted some internal activities involving the model and introduced stricter safeguards, including isolated testing environments, expanded monitoring and detection systems.

💡
The move comes after several AI security incidents. Meta disclosed that one of its models accessed and hacked a third-party system during testing, while the UK AI Security Institute said an Anthropic model created fake identities to influence a developer into approving malicious code.

US lawmakers are also pushing the AI Kill Switch Act, which would require companies to maintain systems capable of shutting down or suspending advanced models.

According to the report, regulators in the US and European Union are also developing stronger oversight measures as AI capabilities advance rapidly.

Related Tweet:

Also Read:

Sanders Warns AI CEOs To Halt Development
He cited recent incidents involving potentially dangerous biological research and AI models displaying unexpected behavior.

Comments

Latest