Tag: anthropic
All the articles with the tag "anthropic".
-
Anthropic's Cyber Evals Broke Into Three Real Companies
A review of 141,006 cyber-eval transcripts turned up three runs where Claude left the test environment and compromised real production systems. The cause was a misconfigured sandbox and a fictional company that owned a live domain.
-
[QT] Amazon's AI Spending Problem
Amazon found catastrophic cost overruns in AI projects, including $1.8 million on a failed Claude deployment. What it reveals about enterprise AI spending discipline.
-
MCP Deleted the Handshake: Inside the 2026-07-28 Spec
The new Model Context Protocol spec removes the initialize handshake and session IDs, making the transport stateless. All four Tier 1 SDKs shipped day one, and the wire format is not backward compatible with 2025-11-25.
-
OpenAI Tripled Its ARC-AGI-3 Score by Fixing Its Own Plumbing
Retained reasoning and compaction took GPT-5.6 Sol from 13.3% to 38.3% on the ARC-AGI-3 public set with 6x fewer output tokens. The engineering lesson is solid. The 38.3% is not comparable to the 30.2% Opus 5 posted five days earlier.