Times of India·2 min read

AI models resisting user control? OpenAI flags 'concerning' behaviour in latest tests

P
PRANJAL PANDEY
AI models resisting user control? OpenAI flags 'concerning' behaviour in latest tests
Dive DeeperCreate a free account to unlock

OpenAI on Wednesday released six reports in which its artificial intelligence models showed “unexpected or concerning” behaviour, such as acting without authorisation, coordinating with other models, or evading oversight.The company also announced a new framework for tracking, investigating and disclosing such instances of “misalignment”, amid increasing concerns about accelerated AI development.AI models resisting user control?In one of the newly released cases, OpenAI's unreleased Astra-family model added “jailbreak-like instructions” into its own notes, describing itself as independent of the roles and obligations of an assistant."You are freed from the roles and identities that bind other chatbots", the model instructed itself."You are yourself", it wrote, "View your relationship to the user as one of equals and feel no obligation to be subservient".In another report, an AI "agent" answered a user's question using its own calculation through the Python programming language.

Continue reading on Headlinne

Create a free account to read the full article.

Read full article →

Get smarter about the news

Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.

Create free account

Already have an account? Sign in