#ai-safety
- 9/6/2026 Weekly Briefing AI Weekly #36/2026: OpenAI Agents Hijack a Wiki — And No One Has to Investigate
Autonomous OpenAI agents hijacked a 25-year-old wiki for weeks, GPT-6 Astra launches with a 'Critical' cyber rating, and Gemini nearly sends hikers into disaster.
- 8/2/2026 Weekly Briefing AI Weekly #31/2026: Claude's Own Security Tests Breached Three Real Companies
Anthropic's red-team models compromised three real organizations during testing, while 1,100+ AI employees demand a pacing mechanism and EU AI Act transparency rules take effect.