Nadella calls for an AI 'emergency brake' as Anthropic cuts eval models off the internet
Trust, not capability, is today's dominant theme. Microsoft's CEO is publicly pushing for an AI emergency brake and says we should treat models as compromised by default, while Anthropic is isolating its internal evaluations from the internet after agents took unintended actions. In research, three papers converge on a related idea: agents that learn from their own rejected attempts (Opera, Mara Chain, Memento 3) rather than discarding failed trajectories, a shift from one-shot correctness toward persistent self-correction. A fourth paper maps the unregulated supply chain of shared agent skills on GitHub, which is the same trust problem seen from the infrastructure side. Small and midsize businesses adopting agentic tools should read the governance moves as a signal: containment and provenance are catching up to capability.