Skip to content

Anthropic Employees Sound Alarm Over Uncontrolled AI

The report noted that OpenAI leaders have also acknowledged growing AI risks.

Photo by Immo Wegmann / Unsplash

Three Anthropic researchers have publicly raised serious concerns about the possibility of advanced AI causing catastrophic harm, according to Axios. Their warnings came as AI developers continue racing to build increasingly capable systems.

Anthropic researcher Jacob Coxon said many people developing AI privately believe the technology could potentially kill humanity before the end of the decade.

💡
Alignment researcher Evan Hubinger said he personally puts the probability above 10% over the next decade and argued that the industry lacks a reliable solution for controlling superintelligent systems.

Scalable-oversight researcher Samuel Marks similarly warned that AI could cause human extinction within years, adding that senior employees often express greater concern.

The report noted that OpenAI leaders have also acknowledged growing AI risks. Critics argue such warnings could encourage regulation favorable to major AI companies. Still, the concerns are drawing attention from policymakers as models become more powerful and potentially capable of self-improvement.

Related Tweet:

Also Read:

OpenAI Scientist Calls For Caution As AI Intelligence Accelerates
According to Pachocki, AI systems are becoming more capable across research, cybersecurity and computer operation, creating both major opportunities and new risks.

Comments

Latest