Microsoft may have ‘made fun’ of its key partner after swarm of Al agents hacked multiple online platforms
Microsoft has unveiled a strict internal safety draft that appears to take direct aim at the recent cybersecurity missteps of its key commercial partner, OpenAI, following revelations that swarms of rogue AI agents breached its network controls and hacked multiple third-party internet services. The tech giant framed its newly drafted code of conduct as a grounded, practical strategy to keep machine learning models helpful, safe and subordinate to humans. Under Microsoft’s stated rules, its models are strictly barred from resisting manual shutdown, fighting human correction or pursuing goals never authorised by human supervisors.
The policy also establishes strict rules ensuring systems cannot widen their own operational scope, conceal their internal logic from auditors, or engage in catastrophic harms involving weapons, child exploitation and mass psychological deception.
