Safety talks go multilateral as Salesforce and labs race on reasoning and agent trust
The day's strongest signal is process, not a model drop: OpenAI, Anthropic, and Google have reportedly been coordinating on AI safety for weeks, while Nvidia's Jensen Huang argues publicly that regulation should stay out of it. That tension sits alongside Salesforce and Nvidia's new reasoning model, which TechCrunch frames as a direct challenge to frontier labs' pricing power, and Ars Technica's finding that paying for frontier models only buys a four-month head start at five times the cost. Research is converging on the same theme from a different angle: a SWE-bench audit shows top coding agents are now statistically indistinguishable, and several papers this week propose "social harnesses" and adversarial stress tests for multi-agent systems that must coordinate or fail safely without constant human oversight. Meanwhile the infrastructure story keeps getting heavier: data centers are on pace to out-consume Germany and Japan's natural gas combined by 2035, and public polling shows persistent unpopularity of both AI and the data centers that power it. Move first, but the industry is starting to reckon with what it costs to stay ahead.