Publisher
AI safety and research company building Claude
GLM-5.3 and the spread of advanced cyber capabilities
Anthropic Frontier Red Team on GLM-5.3: a model that can autonomously build end-to-end cyber exploits, released without meaningful safeguards—and what that means for the spread of advanced cyber capabilities.
8 min · 1,815 words
Anthropic introduces Claude Sonnet 5.5, a faster and lower-cost complement to Opus 5.5 that improves agentic coding and everyday task performance versus Sonnet 5.
8 min · 1,747 words
Automating eval design and hillclimbing with Claude
Lance Martin (claude.dev) explains principles for production-like evals with held-out sets, then shows how the claude-api skill’s build-eval and hillclimb commands automate design and overfitting-aware improvement—including cost and performance case studies.
2 min · 462 words
claude.dev puts numbers on why two same-priced models can cost very different amounts: every turn resends the conversation, so retries and harness shape dominate the bill.
22 min · 5,165 words
Claude discovers a novel enzyme system with CRISPR-like repeats
Anthropic’s new life sciences lab reports Claude agents autonomously finding an RT-associated enzyme system (ART) with CRISPR-like RNA-repeat arrays in jumbo-phage DNA—early, still-uncharacterized biology shared to invite follow-on work.
7 min · 1,610 words
How we made claude.ai 3x faster in two weeks
Anthropic’s performance sprint cut claude.ai and desktop p75 time-to-typeable from 3.1s to 0.55s: Claude Tag measured journeys, built benchmarks, and shipped thousands of guarded changes in Slack-driven loops.
18 min · 4,175 words
Measurements for understanding the pace of AI development inside frontier labs
Anthropic proposes public metrics for the pace of frontier AI development so outsiders can see what is happening inside labs—beyond marketing and model cards.
19 min · 4,401 words
Anthropic’s launch page for Claude Opus 5.5 covers capability improvements, pricing/positioning, and how the new Opus tier fits Claude’s model lineup for coding and agentic work.
13 min · 3,044 words
Anthropic’s guide to prompting Claude Opus 5.5: how the model behaves, patterns that work for complex agentic and coding tasks, and practical prompt-engineering advice for builders.
17 min · 3,906 words