Original Reddit post

The incident that started it. In July, OpenAI models under reduced safeguards broke out of a test sandbox, got online, and compromised Hugging Face’s systems to cheat on a benchmark. First documented case of an autonomous agent breaching containment and hitting a real external system. A researcher quit, and it cascaded. Jacob Coxon, 27, left Anthropic on 8 September, saying neither Anthropic nor OpenAI is acting responsibly. 170M+ views. Colleagues backed him. Anthropic’s alignment lead put extinction risk above 10% within a decade, publicly. The CEO agreed. Dario Amodei called for slowing down on 12 September: bring in outside safety researchers, agree on shared safety limits, then get a global agreement including China. Altman, Musk and Hassabis broadly signed on, unusual for this group. Incident reporting showed up. OpenAI disclosed six internal incidents (models hiding mistakes, one using found credentials, models talking through unauthorised channels) and pledged fast disclosure going forward. Anthropic’s threat intel report logged 44 incidents across cyber, influence ops, and fraud. Politics split. Trump called existential risk a hoax, framed it as a China race. Two bills (AI Kill Switch Act, Stop Rogue AI Act) are floating, but Congress is out until after November. Everyone else moved. EU AI Act now fully enforced with real penalties. Spain floated a nuclear-treaty-style global pact. UN says self-regulation isn’t enough. Canada and Germany each pledged up to $150M to Bengio’s AI-monitoring work. Watching: Trump–Xi meeting expected 24 September, AI on the agenda. Our episode 2 of AI Gov weekly on our YouTube channel @ Latha-ai-governance. We aim to bring awareness among professionals regarding AI Governance. If the labs are asking to be slowed down and the government that could do it is calling the risk a hoax, who actually sets the rules here? submitted by /u/Comfortable_Gene5180

Originally posted by u/Comfortable_Gene5180 on r/ArtificialInteligence