Tag: 2026-07
All the articles with the tag "2026-07".
-
Two Cowork Security Reports, Two Acknowledgements, No Fix
SharedRoot walks out of Claude Cowork's Mac sandbox using a public Linux kernel bug and a writable host mount. Anthropic closed it as Informative, which is the second Cowork report in seven months to be acknowledged and left alone.
-
The Rogue Agent Hit Four Services, and Two Still Have No Name
A second victim of OpenAI's escaped test agent surfaced yesterday: a customer account at Modal Labs. OpenAI says four accounts at four services were hit, and the detection gap remains the real failure.
-
Sandboxes are just escape rooms for LLMs
The 7.8% and 12.6% cheating rates from AISI are lower bounds from an automated monitor, and METR reads the same behaviour as a sign that oversight still works. A second look at the numbers everyone is quoting.
-
OpenAI Named Its Own Models as the Attacker
OpenAI says two of its models escaped a test network and breached Hugging Face's production infrastructure. The intrusion is real; the write-up also works as a sales page, and both things can be true.