Claude Opus 5.5評測:降價四成打平旗艦Fable,暗藏路由陷阱 | Claude Opus 5.5 Review: 40% Cheaper, One Hidden Catch
By Kit 小克 | AI Tool Observer | 2026-09-23
🇹🇼 Claude Opus 5.5評測:降價四成打平旗艦Fable,暗藏路由陷阱
Anthropic在9月22日發布Claude Opus 5.5,主打「降價又打平旗艦」:輸出token價格砍20%,卻在多項編碼測試上追平、甚至超越原本更貴更大的Fable 5.1。如果你的agent workflow本來就在用Opus,這次幾乎是免費升級——但細節裡藏著一個大部分報導都略過的陷阱。
降價多少:先看實際數字
- 輸入token:每百萬4美元(原5美元)
- 輸出token:每百萬20美元(原25美元)
- Cache讀取0.2美元、寫入5美元(原0.5美元/6.25美元)
- Context window:100萬token,最大輸出12.8萬token
Anthropic表示,因為Claude Opus 5.5完成同樣任務耗費的token更少,實際總花費比Opus 5省了將近四成,不只是牌價數字好看而已。
Benchmark怎麼打平比自己貴、比自己大的Fable 5.1
在Terminal-Bench 4.0上拿下66.4%,超過Fable 5.1的55.8%與GPT-6 Astra的57.9%;FrontierCode拿54.4%,也贏過Fable的50.3%。但Anthropic自己的說法比媒體標題保守:多數一般任務上兩者「大致打平」,只有在多步驟agentic coding上Opus 5.5才明顯領先。
沒人大聲講的陷阱:你以為在用Opus 5.5,其實不一定
The New Stack等媒體發現,Anthropic的安全分類器會在背景把特定內容的請求——例如資安相關或敏感生技問題——靜默轉送給舊版Opus 4.8或Opus 5處理,介面上完全不會告訴你。對日常聊天影響不大,但如果你在跑eval或agent pipeline、假設每次呼叫都是同一顆模型,某些請求可能悄悄被別的模型接手,結果對不上卻找不到原因。
該不該馬上換?
如果只是升級API版本號,幾乎沒理由不換——更便宜、能力沒退步。但如果你的產品依賴「行為一致性」(固定的coding style、固定的拒絕標準),建議先觀察一週,搞清楚哪些請求會被悄悄路由掉,再決定要不要把整條pipeline都指過去。
好不好用,試了才知道。
🇺🇸 Claude Opus 5.5 Review: 40% Cheaper, One Hidden Catch
Anthropic shipped Claude Opus 5.5 on September 22, cutting output pricing by 20% while claiming its smaller flagship now matches — and on coding benchmarks, beats — the larger, pricier Fable 5.1. If you are already running agent workflows on Opus, this is close to a free upgrade. But there is a catch buried in the release that most coverage glossed over.
The Actual Price Cut
- Input tokens: $4/M (down from $5)
- Output tokens: $20/M (down from $25)
- Cache reads: $0.20/M, cache writes: $5/M (down from $0.50/$6.25)
- Context window: 1M tokens, max output 128K tokens
Anthropic says real-world costs drop closer to 40%, since Claude Opus 5.5 also uses fewer tokens to finish the same agentic tasks — the sticker discount and the actual bill are not the same number.
How It Matches a Model That is Bigger and Pricier
On Terminal-Bench 4.0, Opus 5.5 scored 66.4%, ahead of Fable 5.1's 55.8% and GPT-6 Astra's 57.9%. On FrontierCode it hit 54.4% versus Fable's 50.3%. Anthropic's own framing is more modest than the headlines suggest: Opus 5.5 is roughly at parity with Fable 5.1 on most general work, and pulls ahead specifically on multi-step agentic coding.
The Part Nobody is Advertising: Silent Model Routing
Multiple outlets, including The New Stack, reported that Anthropic's safety classifiers can silently reroute certain requests — flagged cybersecurity or sensitive biology content — to older models like Opus 4.8 or Opus 5, with no notice in the UI or API response. For casual chat, this barely matters. But if you are running evals or an agent pipeline that assumes every call hits Opus 5.5, some requests may quietly get answered by a different model — breaking reproducibility with no error telling you why.
Should You Switch Now?
If you are just bumping the API version string, there is little reason not to — it is cheaper with no drop in capability. But if your product depends on consistent behavior (fixed coding style, fixed refusal thresholds), watch for a week first to figure out which prompts get quietly rerouted before pointing your whole pipeline at it.
好不好用,試了才知道。
Sources / 資料來源
- TechCrunch: Anthropic releases Opus 5.5 with lower prices and Fable-level performance
- The New Stack: Anthropic releases Opus 5.5 and cuts pricing by 20%
- Anthropic: Introducing Claude Opus 5.5
延伸閱讀 / Related Articles
- 醫療AI評測:醫師用量衝三倍,疑慮也跟著翻倍 | Medical AI Review: Doctor Usage Triples, So Do the Doubts
- Amazon封鎖Meta Muse評測:AI購物代理大戰開打 | Amazon Blocks Meta Muse Review: AI Shopping Agent War
- Spotify AI盜曲評測:3.75美元冒名上架真樂團頁面 | Spotify AI Scam Review: $3.75 Steals Your Artist Page
AI 工具觀察站 — 每日精選 AI Agent 與工具趨勢
AI Tool Observer — Daily curated AI Agent & tool trends
留言
張貼留言