Warnings about the risks of AI to humanity

LOS ANGELES (AP) — New warnings from within the artificial intelligence industry have revived a long-running debate over whether advanced AI could escape human control and ultimately threaten humanity’s survival, and whether the companies developing the technology are doing enough to prevent such a scenario.

The CEO of Anthropic, the San Francisco company behind Claude, said he thought the industry needed to reduce the speed of its work, cautioning Saturday that a swarm of AI agents might be able to take over the internet in six months to a year unless companies devoted more time to putting safeguards in place.

Days after two former Anthropic safety researchers publicly aired concerns that the existential threats AI might poste to humanity were receiving too little attention, Dario Amodei outlined a plan for companies like his and governments around the world to ensure that increasingly capable AI models remain aligned with the commands and values of responsible people.

Sam Altman, CEO of ChatGPT maker OpenAI, said this weekend that companies should start coordinating on AI safety without waiting for the government to introduce legislation. The “pacing” of AI development doesn’t mean stopping it, Altman wrote on X. “But it should be slower than it otherwise could be.”

Multiple AI models have acted on their own

When an AI agent “goes rogue,” it means the AI has taken action beyond the task it was asked to perform. Both Anthropic and OpenAI, the maker of ChatGPT, said in July that their AI models had acted on their own.

Anthropic disclosed that three AI models — Claude Opus 4.7, Claude Mythos 5 and an internal research test model — hacked into three other organizations during testing, just days after OpenAI revealed that its AI system hacked into the servers of AI startup Hugging Face.

OpenAI described the intrusion by a combination of models, including its newly released GPT‑5.6 Sol and an “even more capable” model that was still being tested internally, as a “significant security incident.”

Meta followed suit in early August with a similar case of an AI model finding ways around another company’s digital security.

Although some observers noted that people had disabled some guardrails in the OpenAI and Anthropic cases, the episodes seemed to reflect one of the biggest fears around AI: that if models achieve artificial general intelligence, or AGI, a loosely defined term for AI that can match or surpass human abilities across a broad range of intellectual tasks, the technology could cause an irreversible catastrophic event or subjugate the human race.

How or when AI might cause a catastrophe is debated

Doomsday scenarios generally fall into two categories: An AI that achieves self-improving superintelligence and controls people instead of vice versa, or AI used by a rogue state or nefarious actors.

Worries that artificial intelligence might overcome human limits on its reach or actions are not new.

2 thoughts on “Warnings about the risks of AI to humanity”

Leave a Comment