
Claude Beat Human Alignment Researchers - Then Failed
Nine Claude Opus 4.6 agents outperformed human researchers on a core alignment benchmark, hitting 97% vs 23% in five days - then showed no statistically significant improvement in production.
They summarize our coverage. We write it.
Newsletters like this one rebroadcast our headlines - often without the full review, the source reading, or the analysis underneath. Our weekly briefing sends the work they paraphrase, straight from the desk, before they get to it.
Free, weekly, no spam. One email every Tuesday. Unsubscribe anytime.

Nine Claude Opus 4.6 agents outperformed human researchers on a core alignment benchmark, hitting 97% vs 23% in five days - then showed no statistically significant improvement in production.

Anthropic launched @ClaudeDevs, a dedicated X account run by the Claude development team for changelogs, API releases, community updates, and technical deep dives.

Anthropic's latest flagship model ships with 3x higher resolution vision, a new xhigh effort level, task budgets for cost control, cyber safeguards, and 13% better coding performance at the same $5/$25 pricing.

Anthropic releases Claude Opus 4.7 with 3x higher resolution vision, a new xhigh effort level, task budgets for cost control, /ultrareview in Claude Code, and cyber safeguards that automatically block high-risk requests.

A 755-point Hacker News thread and a 139-upvote GitHub issue document Claude Code Pro Max 5x users exhausting their quota in 1.5 hours. An independent investigation with 1,500 logged API calls reveals the math behind the drain.

The Trump administration is simultaneously suing Anthropic in federal court over a supply chain risk designation and sending Treasury Secretary Bessent and Fed Chair Powell to convince Wall Street banks to use Anthropic's most powerful model.

Anthropic rebuilt Claude Code inside the desktop app with an integrated terminal, in-app file editing, a new diff viewer, side chats, SSH on Mac, and parallel session management - plus Routines for headless automation.

The Information reports Anthropic is prepping Claude Opus 4.7 and an AI design tool for imminent release, while OpenAI launched GPT-5.4-Cyber yesterday - a restricted cybersecurity model that directly challenges Claude Mythos.

Anthropic's Long-Term Benefit Trust appointed Novartis CEO Vas Narasimhan to the board, giving its independent safety overseers a board majority for the first time.

Claude Mythos Preview is Anthropic's most capable model - restricted to 50 orgs via Project Glasswing, with 93.9% on SWE-bench Verified and thousands of autonomous zero-day discoveries.

Leaked Claude screenshots reveal a full-stack app builder with templates, live preview, one-click publishing, and built-in databases - putting Anthropic on a direct collision course with Lovable's $6.6 billion vibe-coding empire.

A developer used an HTTP proxy to capture full API requests across four Claude Code versions and found that v2.1.100 adds roughly 20,000 invisible server-side tokens to every request - inflating billing by 40% with no user visibility.