Skip to content

Anthropic

Topic archive • 3 matches

Back to home • GEO summary endpoint

2026-09-29

Technology

  • Sonnet 5.5 Is Here. Look What It Can Build.: Anthropic's Claude 3.5 Sonnet model demonstrates its coding capabilities by building several interactive web projects and games. The showcased projects include Crownfall, Bounce Lab, Marrow Manor, Splatburst, and Deadlock. The demonstration also compares these results against previous generations of the Sonnet model.

    Technology • Matthew Berman

    Permalink
  • Anthropic launches Claude Sonnet 5.5 with 30% lower costs per task: Anthropic has launched Claude Sonnet 5.5, which generates output over 30 percent faster and costs up to 30 percent less per task than its predecessor. The model nearly matches Opus 5.5 on knowledge-work benchmarks and improved its score on the Terminal-Bench coding benchmark from 10.3 to 70.6 percent.

    Anthropic • The Decoder

    Permalink
  • Anthropic develops Claude Code workflow to prevent false benchmark gains: Anthropic has developed a new workflow for Claude Code designed to build real-world evaluations and hillclimb AI agents against them. The system is built to reject benchmark gains that fail to generalize to unseen tasks, preventing false improvements.

    AI Agents • The Neuron

    Permalink