CNBC·3 min read·medium

AI self-improvement fears prompt ‘existential’ concerns at Anthropic, OpenAI

K
Kai Nicol-Schwarz
AI self-improvement fears prompt ‘existential’ concerns at Anthropic, OpenAI
AI Summary

Researchers at Anthropic and OpenAI are expressing growing concerns regarding the risks of recursive self-improvement in AI models. The fear is that as AI systems become capable of autonomously improving their own code, humans may lose control over their development.

Why it matters

The debate highlights the tension between rapid AI capability advancement and the safety protocols required to prevent existential risks.

Dive DeeperCreate a free account to unlock

This report is from this week's The Tech Download newsletter. Like what you see? You can subscribe here.

Fears over the safety of AI systems — and their potential to wipe out humanity — gained new, viral traction this week.

Evan Hubinger, an alignment lead at Anthropic, said on X that he thinks there is more than a 10% chance that AI could kill all humans within the next decade, after a colleague quit over safety fears.

More warnings from researchers at both Anthropic and OpenAI followed. Cue a social media frenzy.

But it was in Hubinger's reply to his own post that revealed where exactly his concerns lay.

"What I am worried about is superintelligence arising from recursive self-improvement, as we have said is happening faster than we thought," he said.

Continue reading on Headlinne

Create a free account to read the full article.

Read full article →
technologyai

Get smarter about the news

Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.

Create free account

Already have an account? Sign in