Claude Opus 5.5 High ranks second on Agent Arena with 64% cost reduction
What happened
Claude Opus 5.5 (High) 在 Agent Arena 榜单上拿了第 2 名,净改进分 +12.15%,只输给 Fable 5.1 (Max)。官方说它比上一代 Opus 5 (Max) 便宜 56%,但正文没给出具体价格和延迟数据,这个省钱幅度我先打个折,等实际跑量再看。
From AI HOT 精选
Coverage
Follow the reports to see the story from different sides.
- AI HOT (Curated Pool)PickClaude Opus 5.5 (High) hits #2 on Agent Arena, undercuts peers by 64% on cost
Claude Opus 5.5 (High) landed #2 on Agent Arena with a +12.15% net gain, behind only Claude Fable 5.1 (Max). Median cost per task is $1.31—64% cheaper than peers at the same tier, 40% below Opus 5 (High), and 56% below Opus 5 (Max). It ranked #1 on Steerability at +14.50%. The post doesn't break down task mix or latency.
- AI HOT (Curated Pool)PickClaude Opus 5.5 (High) hits #2 on Agent Arena and reshapes the Pareto frontier
Anthropic's Claude Opus 5.5 (High) landed at #2 on Agent Arena with a +12.15% net improvement, behind only Fable 5.1 (Max). Median cost is $1.31 per task—40% cheaper than Opus 5 (High) and 56% cheaper than Opus 5 (Max). It ranked #1 on Steerability at +14.50%. The post doesn't disclose a release date or other model comparisons.
- AI HOT (Curated Pool)PickClaude Opus 5.5 (High) hits #2 on Agent Arena, costs 56% less than Opus 5 (Max)
Anthropic's Claude Opus 5.5 (High) reached #2 on Agent Arena with a +12.15% net improvement, behind only Fable 5.1 (Max). It also costs 56% less than Opus 5 (Max). The post doesn't disclose exact pricing or latency—I'd discount the cost claim until we see real usage numbers.
Heat over time
Not enough continuous observations to draw a trend yet.