Article may be outdated

This article is 2 days old. Some details may have changed since publication.

TechCrunch·3 min read·hard

OpenAI’s new reasoning technique alarms AI safety experts

R
Russell Brandom
OpenAI’s new reasoning technique alarms AI safety experts
AI Summary

AI safety experts are expressing concern over OpenAI's new 'opaque recurrence' reasoning technique in its Astra model. Critics argue that this method makes the model's decision-making process less transparent and harder to monitor for safety.

Why it matters

The shift toward less interpretable AI reasoning models poses a challenge to current safety and alignment verification standards.

Dive DeeperCreate a free account to unlock

OpenAI’s new Astra model will use a reasoning technique called “recurrent depth” that allows it to operate outside of the sequential thinking that characterizes most reasoning models, the The Information reported on Tuesday. This technique, also called “opaque recurrence,” will likely make the model’s chain of thought more difficult to monitor — and that has AI safety experts rattled.

While Astra’s use of the technique is reportedly limited, its emergence has still raised significant concerns among AI safety experts.

“I am extremely concerned by the reporting that Astra uses opaque recurrence,” wrote Redwood CEO Buck Shlegeris in a post after the news broke. “I don’t know whether Astra is much less CoT monitorable than previous models. But if OpenAI pushes this technique further, they’ll have the option to massively increase the recurrence and totally destroys CoT monitorability.”

Continue reading on Headlinne

Create a free account to read the full article.

Read full article →
technologyai

Get smarter about the news

Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.

Create free account

Already have an account? Sign in