In a Saturday post, Nadella said companies should treat powerful AI models as potential insider threats, assume they're already compromised, and build an emergency brake humans control. He wants the whole trust architecture rethought.
Take: When the CEO himself starts yelling about brakes, the internal evals are worse than anyone's admitting. But who holds the brake? Model makers policing themselves is the fox designing the henhouse door.