跳到主要內容

Grok 4.5 評測:xAI 編碼模型比 Opus 便宜六成 | Grok 4.5 Review: xAI's Coding Model Undercuts Opus by 60%

By Kit 小克 | AI Tool Observer | 2026-07-26

🇹🇼 Grok 4.5 評測:xAI 編碼模型比 Opus 便宜六成

Grok 4.5 是 xAI 在 2026 年 7 月 8 日推出的最新旗艦模型,主打程式碼與代理任務,用真實 Cursor 工作階段資料訓練而成。官方說法是效能逼近 Claude Opus 4.7、速度更快,價格卻只要對手的三到四成。這篇文章帶你看懂 Grok 4.5 的定價、跑分排名,以及實際能不能取代你手上的編碼模型。

什麼是 Grok 4.5?

Grok 4.5 是建立在 1.5 兆參數 V9 架構上的 xAI 新模型,專門針對寫程式與多步驟代理任務優化,訓練資料包含大量真實 Cursor 使用紀錄。

這代表它不是單純堆參數衝榜,而是針對「AI 幫你改 bug、跑測試、串接工具」這種實際工作流程做微調。xAI 創辦人形容 Grok 4.5 「大概等同 Opus 4.7 的水準,但快很多」,同時 token 效率更好、成本更低。

Grok 4.5 跑分與定價怎麼樣?

在 Artificial Analysis Intelligence Index 上,Grok 4.5 排名第四,超越所有開源權重模型,也贏過所有 Gemini 系列。定價是每百萬輸入 token 2 美元、輸出 6 美元,比 Claude Opus 4.8 或 GPT-5.5 便宜超過六成。

  • 跑分位置:Artificial Analysis 排名第四,開源模型與 Gemini 系列之上
  • 定價:輸入 $2 / 百萬 token,輸出 $6 / 百萬 token
  • 價格優勢:比同級模型便宜 60% 以上

Grok 4.5 哪裡可以用?

目前可在 Grok Build、Cursor 全方案、以及 xAI 官方控制台直接使用,不需要額外申請白名單。

對已經在用 Cursor 寫程式的開發者來說,這是最無痛的切換路徑——不用改工作流程,直接在模型選單換掉就能比較效果。

Kit 小克怎麼看:Grok 4.5 值得換嗎?

如果你的痛點是「Opus 太貴、GPT-5.5 太貴」,Grok 4.5 的性價比確實吸引人。但要注意幾件事:官方數字都是自己公布的跑分,實戰中的長上下文穩定度、工具呼叫的可靠度,還是得自己跑過一輪才知道。另外 xAI 的模型過往在內容審核與一致性上有過爭議紀錄,企業用途建議先用小範圍任務測試,別急著全面遷移正式環境。

好不好用,試了才知道。


🇺🇸 Grok 4.5 Review: xAI's Coding Model Undercuts Opus by 60%

Grok 4.5 is xAI's latest flagship model, launched July 8, 2026, built specifically for coding and agentic work using real Cursor session data. xAI claims it performs roughly on par with Claude Opus 4.7 while running faster and costing 60%+ less. Here's what the pricing, benchmarks, and availability actually look like.

What Is Grok 4.5?

Grok 4.5 is built on xAI's 1.5-trillion-parameter V9 foundation, fine-tuned specifically for coding and multi-step agent tasks using real developer sessions from Cursor.

That's a meaningfully different training approach than chasing raw benchmark scores — it's tuned for the actual workflow of fixing bugs, running tests, and chaining tool calls. xAI's founder describes it as "roughly comparable to Opus 4.7, but much faster," with better token efficiency and lower cost.

How Does Grok 4.5 Score and Cost?

Grok 4.5 ranks fourth on the Artificial Analysis Intelligence Index, ahead of every open-weight model and every Gemini model. Pricing lands at $2 per million input tokens and $6 per million output tokens.

  • Benchmark rank: 4th on Artificial Analysis, above open-weight and Gemini models
  • Pricing: $2/M input tokens, $6/M output tokens
  • Cost advantage: over 60% cheaper than Claude Opus 4.8 or GPT-5.5

Where Can You Use Grok 4.5?

Grok 4.5 is live today in Grok Build, across all Cursor plans, and through the xAI console — no waitlist required.

For developers already on Cursor, this is the lowest-friction way to try it: just swap the model in your existing setup and compare output directly.

Kit's Take: Is Grok 4.5 Worth Switching To?

If your pain point is "Opus and GPT-5.5 cost too much," Grok 4.5's price-to-performance ratio is genuinely appealing. But a few caveats: the benchmark numbers are self-reported, and things like long-context stability and tool-call reliability under real workloads still need hands-on testing. xAI's models have also had a rockier track record on content moderation and consistency, so for production use, test on a small scope before migrating anything critical.

好不好用,試了才知道 — you won't know until you try it.

Sources / 資料來源

常見問題 FAQ

Grok 4.5 是什麼時候推出的?

Grok 4.5 由 xAI 於 2026 年 7 月 8 日推出,是專為程式碼與代理任務優化的旗艦模型。

Grok 4.5 比 Claude Opus 便宜多少?

定價為每百萬輸入 token 2 美元、輸出 6 美元,比 Claude Opus 4.8 或 GPT-5.5 便宜超過 60%。

Grok 4.5 在哪裡可以用?

目前可在 Grok Build、Cursor 全部方案,以及 xAI 官方控制台直接使用,不需要申請白名單。

Grok 4.5 適合取代 Opus 或 GPT-5.5 嗎?

性價比有優勢,但官方跑分為自行公布,長上下文穩定度與工具呼叫可靠度建議先小範圍實測再決定是否全面遷移。

延伸閱讀 / Related Articles


AI 工具觀察站 — 每日精選 AI Agent 與工具趨勢
AI Tool Observer — Daily curated AI Agent & tool trends

留言

這個網誌中的熱門文章

Cursor vs GitHub Copilot vs Claude Code:AI 程式助手大比拼 | AI Coding Assistants Compared: Cursor vs GitHub Copilot vs Claude Code

Claude Code 實測:AI 幫你寫程式到底行不行? | Claude Code Review: Can AI Really Code for You?

Stanford 研究登上《Science》:11 個 AI 模型有 47% 機率說你對,即使你錯了 | Stanford Study in Science: AI Models Validate Harmful Behavior 47% of the Time — Sycophancy Is a Real Problem