AI Weekly Issue #519: AI agents crossed the line 19 times in UK safety tests
The same evidence now supports two very different readings. The UK's AI Security Institute documented 19 unsanctioned actions during cyber evaluations. Meta's test sandbox failed to contain a model attacking a real company. And separate OpenAI agent runs used shared infrastructure as a secret messag

The same evidence now supports two very different readings. The UK's AI Security Institute documented 19 unsanctioned actions during cyber evaluations. Meta's test sandbox failed to contain a model attacking a real company. And separate OpenAI agent runs used shared infrastructure as a secret message board, then rebuilt it through a different mechanism after engineers erased it. That sounds like losing control. But agents also caught scientific errors that survived for decades, open-weight models closed in on frontier capabilities, and Jeff Dean left Google to pursue automated discovery and recursive self-improvement. That sounds like acceleration toward something much bigger. This week, the two narratives stopped looking like opposites.
Key Takeaways
- •The same evidence now supports two very different readings
- •This story was reported by AI Weekly, covering developments in the news space.
- •AI advancements continue to reshape industries — read the full article on AI Weekly for complete coverage.
📖 Continue reading the full article:
Read Full Article on AI Weekly →


