---
format: "aidr-story-markdown/v1"
id: "54a584b50c10ead2634337a6de773b6a3326dd1916a9f5fec6fdecb1eddc4e07"
canonical_url: "https://aidr.today/54a584b5?lang=en"
title: "Claude Opus 5.5 Debuts at No. 2 in Agent Arena at 64% Lower Cost Per Task"
lang: "en"
requested_lang: "en"
available_langs: ["en","vi"]
translation_fallback: null
fallback_fields: []
published_at: "2026-09-28T17:10:30.000Z"
category: "Releases"
topics: ["anthropic","llm","agent"]
source_urls: ["https://huggingnews.com/ai/update-claude-opus-55-debuts-at-no-2-in-agent-arena-at-64percent-lower-c-6b148e30","https://x.com/martinvars/status/2102466916874887577","https://marketbrief.now/ai/update-claude-opus-55-debuts-at-no-2-in-agent-arena-at-64percent-lower-c-6b148e30","https://huggingnews.com/ai/cognition-cuts-devin-costs-up-to-70percent-and-tops-frontiercode-11-2e066e80","https://marketbrief.now/ai/cognition-cuts-devin-costs-up-to-70percent-and-tops-frontiercode-11-2e066e80"]
summary: "Claude Opus 5.5 debuted at No. 2 in Agent Arena with a net improvement score of +12.15%, behind only Fable 5.1 (Max), while its $1.31 median price per task came in at 64% less cost and pushed out the Pareto frontier. The High setting posts a higher net improvement score than both prior Opus 5 variants while costing 40% less than Opus 5 (High) and 56% less than Opus 5 (Max). By signal, it ranks first in steerability (+14.50%), second in confirmed success (+15.50%), third in praise versus complaint (+19.80%) and fourth in bash recovery (+10.64%). The result extends a run of leaderboard entries for the model, which also ranks No. 1 on Xbench with +76. A faster Sonnet 5.5 is being grey-tested in Claude Code at $2 per million input tokens and $10 per million output tokens, with cache reads at $0.20. Anthropic released Opus 5.5 on Sept. 22 as the first model in its Claude 5.5 family, priced at $4/$20 per million input/output tokens, down 20% from Opus 5. The company says it matches Fable 5.1 on most tasks and costs 40% less to run than Opus 5 on typical workloads, with output generated more than 30% faster. Sonnet 5.5 and Haiku 5.5 are due in the coming weeks."
---

# Claude Opus 5\.5 Debuts at No\. 2 in Agent Arena at 64% Lower Cost Per Task

> [Open the canonical story](<https://aidr.today/54a584b5?lang=en>)

**Published:** 2026-09-28T17:10:30.000Z
**Category:** Releases
**Topics:** anthropic, llm, agent

## Summary

Claude Opus 5\.5 debuted at No\. 2 in Agent Arena with a net improvement score of \+12\.15%, behind only Fable 5\.1 \(Max\), while its $1\.31 median price per task came in at 64% less cost and pushed out the Pareto frontier\. The High setting posts a higher net improvement score than both prior Opus 5 variants while costing 40% less than Opus 5 \(High\) and 56% less than Opus 5 \(Max\)\. By signal, it ranks first in steerability \(\+14\.50%\), second in confirmed success \(\+15\.50%\), third in praise versus complaint \(\+19\.80%\) and fourth in bash recovery \(\+10\.64%\)\. The result extends a run of leaderboard entries for the model, which also ranks No\. 1 on Xbench with \+76\. A faster Sonnet 5\.5 is being grey\-tested in Claude Code at $2 per million input tokens and $10 per million output tokens, with cache reads at $0\.20\. Anthropic released Opus 5\.5 on Sept\. 22 as the first model in its Claude 5\.5 family, priced at $4/$20 per million input/output tokens, down 20% from Opus 5\. The company says it matches Fable 5\.1 on most tasks and costs 40% less to run than Opus 5 on typical workloads, with output generated more than 30% faster\. Sonnet 5\.5 and Haiku 5\.5 are due in the coming weeks\.

## Sources

- [Story source](<https://huggingnews.com/ai/update-claude-opus-55-debuts-at-no-2-in-agent-arena-at-64percent-lower-c-6b148e30>)
- [Story source](<https://x.com/martinvars/status/2102466916874887577>)
- [Story source](<https://marketbrief.now/ai/update-claude-opus-55-debuts-at-no-2-in-agent-arena-at-64percent-lower-c-6b148e30>)
- [Story source](<https://huggingnews.com/ai/cognition-cuts-devin-costs-up-to-70percent-and-tops-frontiercode-11-2e066e80>)
- [Story source](<https://marketbrief.now/ai/cognition-cuts-devin-costs-up-to-70percent-and-tops-frontiercode-11-2e066e80>)

