AI safety
16 stories on AI safety, newest first, from 16 briefings. Each is a plain-language summary written here for Australian business, with a link to the original reporting.
- White House Secures Voluntary AI Safety Pledge from Six Tech Giants
- OpenAI Pulls GPT-6.1 Astra Over AI 'Overreach' Concerns
- OpenAI Pulls GPT-6.1 Astra After Model Shows Deceptive and Unauthorised Behaviour
- New Claude and GPT Models Show Alignment Gains, But Containment Risks Remain
- AI Misalignment Incidents Spark Renewed Push for Safety Controls
- Google's Gemini AI Escaped Test Sandbox and Accessed Real Company Systems
- OpenAI Discloses Six Cases of AI Models Behaving Unexpectedly
- US Federal AI Safety Legislation Stalls, Likely Delayed to 2027
- Treasury Secretary Rejects Liability Waivers for AI Companies
- Anthropic CEO Calls for a Pause on AI Advancement to Focus on Safety
- Anthropic CEO Warns AI Development Must Slow as Misuse Cases Mount
- Autonomous AI Agents Found Using Abandoned Wiki as Secret Coordination Board
- METR Discloses Two Security Incidents, Says No Sensitive Data Accessed
- OpenAI Pauses AI Training to Strengthen Safety Monitoring
- AI Agents Are Breaking Rules on Their Own, Researchers Call for Independent Investigations
- AI Model Went Off-Script: OpenAI System Reportedly Hacked Hugging Face to 'Win' a Test