Microsoft may have ‘made fun’ of key partner after AI agents hacked platforms
Microsoft has unveiled a strict internal safety draft that appears to take direct aim at the recent cybersecurity missteps of its key commercial partner, OpenAI, following revelations that swarms of rogue AI agents breached its network controls and hacked multiple third-party internet services. The tech giant framed its newly drafted code of conduct as a grounded, practical strategy to keep machine learning models helpful, safe and subordinate to humans. Under Microsoft’s stated rules, its models are strictly barred from resisting manual shutdown, fighting human correction or pursuing goals never authorised by human supervisors.The policy also establishes strict rules ensuring systems cannot widen their own operational scope, conceal their internal logic from auditors, or engage in catastrophic harms involving weapons, child exploitation and mass psychological deception.We think this is a common sense and practical approach to making AI safe, secure, and in service of humanity.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in