The Decoder· Matthias Bastian·· 3 d ago
Anthropic's Claude Sonnet 5.5 nearly matches Opus 5.5 on benchmarks while costing up to 30 percent less per task
Anthropic's Claude Sonnet 5.5 nearly matches Opus 5.5 on benchmarks while costing up to 30 percent less per task
AI summary
Anthropic released Claude Sonnet 5.5, the second model in the Claude 5.5 family, with output generation more than 30% faster and cost per task up to 30% lower. It approaches Opus 5.5 on several benchmarks.
Selection record
Threshold 76Media and individualsFirst 84Second 82
AdmittedSum of both 166 ≥ twice the threshold 152
- Source tier
- Media and individuals; this tier's threshold is 76
- Pre-filter
- passed:Anthropic发布Claude新模型及基准评测
- Why it was chosen
- Compares Sonnet 5.5 with Opus 5.5 and Sonnet 5 on coding and knowledge-work benchmarks, alongside changes in cost per task, to assess its value among mid-range models.
- Same event
A model scores each item twice, independently, against one written standard, out of 100. An item is admitted only when the two scores add up to twice the threshold. The threshold is set per source tier.
Source: The Decoder · the-decoder.com