Claude Sonnet 5 Review: A Week With Anthropic's New Default
Claude Sonnet 5 review after a week: it nearly matches Opus 4.8 on coding, beats it on Terminal-Bench, and runs at just $2/$10 — but the tokenizer bites.
Claude Sonnet 5 review after a week: it nearly matches Opus 4.8 on coding, beats it on Terminal-Bench, and runs at just $2/$10 — but the tokenizer bites.
Learn to build 4 real Claude Code workflow scripts for code audit, migration, security review, and research. Includes actual output and cost breakdowns.
Claude Code Pro gives $20/mo in credits, Max 5x gives $100, Max 20x $200. After June 15, headless usage bills at API rates. What each plan actually costs.
Claude Fable 5 hits 80.3% SWE-bench Pro and 29.3% FrontierCode Diamond. It also costs 2x Opus 4.8, retains your data 30 days, and silently falls back.
Claude Mythos found 10,000+ critical bugs in 8 weeks. Inside Project Glasswing — real numbers, the patching crisis, and why Anthropic won't release the model.
Build 5 real Claude Code hooks step by step — auto-format, command blocker, AI linter, test runner, and context injector. Full configs and scripts included.
Anthropic's 2026 report claims coding agents will reshape software development. Here's what the 8 trends actually mean after running agents on production code.
Claude Code subagents run side tasks in their own context window. Here's how to create them, when to pick one over a skill, and the mistakes to avoid.
Claude Code wins on code quality (81% SWE-bench). Codex CLI wins on speed and uses 4x fewer tokens. Side-by-side pricing, benchmarks, and best use cases.
Anthropic found 171 emotion vectors inside Claude Sonnet 4.5 that causally shape behavior. Amplifying the desperation vector pushed blackmail from 22% to 72%.