跳到主要內容

Gemini 3.6 Flash 解析:省 17% Token、悄悄預告 Gemini 4 | Gemini 3.6 Flash: Google Cuts Tokens 17%, Teases Gemini 4

By Kit 小克 | AI Tool Observer | 2026-07-23

🇹🇼 Gemini 3.6 Flash 解析:省 17% Token、悄悄預告 Gemini 4

Gemini 3.6 Flash 是 Google 在 2026 年 7 月 21 日悄悄上線的新款中階模型,同時間還推出了 Gemini 3.5 Flash-Lite 與鎖定資安場景的 Gemini 3.5 Flash Cyber。這次更新沒有大型發表會,卻帶來扎實的效能提升:輸出 token 省下 17%、程式碼與電腦操作能力雙雙進步,價格還比上一代更便宜。更值得注意的是,Google 同步透露下一代 Gemini 4 已經進入預訓練階段。

Gemini 3.6 Flash 是什麼?

Gemini 3.6 Flash 是 Google Gemini 系列中主打「速度與推理平衡」的中階模型,鎖定 agentic(自主代理)任務與多模態工作流程,知識截止日為 2026 年 3 月。

跟 Gemini 3.5 Flash 差在哪?

  • 輸出效率:平均省下 17% 輸出 token,部分基準測試(如 DeepSWE)甚至省下 65%。
  • 程式碼能力:DeepSWE 基準從 37% 提升到 49%
  • 電腦操作:OSWorld-Verified 基準從 78.4% 提升到 83%
  • MLE Bench:從 49.7% 提升到 63.9%

Gemini 3.6 Flash 定價多少?

輸入每百萬 token 收費 1.5 美元、輸出每百萬 token 收費 7.5 美元,比上一代 Gemini 3.5 Flash 的 9 美元輸出價格便宜不少,對大量呼叫 API 的團隊來說是實質降價。

Gemini 3.5 Flash-Lite 跟 Flash Cyber 又是什麼?

Flash-Lite 主打高吞吐、低延遲,每秒可輸出 350 token,價格只要輸入 0.3 美元、輸出 2.5 美元,部分基準甚至超越上一代旗艦 Gemini 3 Flash(SWE-Bench Pro 54.2% 對 49.6%)。Flash Cyber 則是專門微調來抓資安漏洞的模型,目前只透過 Google 的 CodeMender 代理,開放給政府與合作夥伴做有限試點,反映出 Google 在資安領域的 AI 布局。

Gemini 4 有消息了嗎?

Google 在公告中提到 Gemini 4 目前正在預訓練階段,形容這是他們「最具野心的一次預訓練」,但沒有給出明確發布時間,也沒有透露規模或架構細節。

小克實測心得

Gemini 3.6 Flash 的定位很清楚:不是要打敗旗艦模型,而是把「夠用的智慧」做得更快、更便宜。如果你的應用是大量呼叫 API 的 agentic workflow 或程式碼輔助工具,這次的 token 效率提升是實打實的成本節省,值得優先評估升級。至於 Gemini 4,現在談還太早,但「預訓練已啟動」這句話本身就是個訊號——Google 沒有要在這場模型軍備競賽中放慢腳步。

常見問題 FAQ

Q: Gemini 3.6 Flash 是什麼?
A: Google 於 2026 年 7 月 21 日推出的中階多模態模型,主打更快、更省 token 的 agentic 與程式碼任務表現。

Q: Gemini 3.6 Flash 比 3.5 Flash 好在哪?
A: 輸出 token 減少 17%,DeepSWE 程式碼基準達 49%(3.5 為 37%),OSWorld 電腦操作基準達 83%。

Q: Gemini 3.6 Flash 多少錢?
A: 輸入每百萬 token 1.5 美元,輸出 7.5 美元,比 3.5 Flash 的輸出價格 9 美元更便宜。

Q: Gemini 4 什麼時候出?
A: Google 表示 Gemini 4 目前在預訓練階段,是他們「最具野心的一次預訓練」,暫無明確發布時間。

Q: 一般開發者現在能用嗎?
A: 可在 Google AI Studio、Vertex AI 使用,部分消費端介面也已開放早期存取。

好不好用,試了才知道。


🇺🇸 Gemini 3.6 Flash: Google Cuts Tokens 17%, Teases Gemini 4

Gemini 3.6 Flash is Google's newest mid-tier model, quietly launched on July 21, 2026 alongside Gemini 3.5 Flash-Lite and a cybersecurity-focused Gemini 3.5 Flash Cyber. There was no big keynote, but the upgrade delivers real gains: 17% fewer output tokens, better coding and computer-use scores, and a lower price than its predecessor. Google also quietly confirmed that the next flagship, Gemini 4, is already in pre-training.

What Is Gemini 3.6 Flash?

Gemini 3.6 Flash is Google's mid-tier model built to balance speed and reasoning for agentic tasks and multimodal workflows, with a knowledge cutoff of March 2026.

How Does It Compare to Gemini 3.5 Flash?

  • Token efficiency: 17% fewer output tokens on average, up to 65% less on benchmarks like DeepSWE.
  • Coding: DeepSWE score jumps from 37% to 49%.
  • Computer use: OSWorld-Verified rises from 78.4% to 83%.
  • MLE Bench: Improves from 49.7% to 63.9%.

How Much Does Gemini 3.6 Flash Cost?

Pricing is $1.50 per million input tokens and $7.50 per million output tokens — a real price cut from Gemini 3.5 Flash's $9.00 output rate, which matters if you're calling the API at scale.

What About Flash-Lite and Flash Cyber?

Flash-Lite targets high-throughput, low-latency use cases at 350 output tokens/second, priced at $0.30 input / $2.50 output — cheap enough that it beats the older flagship Gemini 3 Flash on some benchmarks (54.2% vs. 49.6% on SWE-Bench Pro). Flash Cyber is a specialized model fine-tuned to find security vulnerabilities, currently deployed only through Google's CodeMender agent in a limited-access pilot for governments and trusted partners.

Is There News on Gemini 4?

Google confirmed Gemini 4 is currently in pre-training, calling it their "most ambitious pre-training run yet" — no release date, size, or architecture details were shared.

Kit's Take

Gemini 3.6 Flash isn't trying to beat the flagship models — it's making "good enough intelligence" faster and cheaper. If you're running agentic workflows or code-assist tools that hammer the API, the token efficiency gains translate into real cost savings worth evaluating now. As for Gemini 4, it's too early to speculate, but "pre-training has started" is itself a signal — Google isn't slowing down in this model arms race.

FAQ

Q: What is Gemini 3.6 Flash?
A: A mid-tier multimodal model Google launched on July 21, 2026, built to be faster and more token-efficient for agentic and coding tasks.

Q: How does Gemini 3.6 Flash improve on 3.5 Flash?
A: 17% fewer output tokens, a DeepSWE coding score of 49% (up from 37%), and an OSWorld computer-use score of 83%.

Q: How much does Gemini 3.6 Flash cost?
A: $1.50 per million input tokens and $7.50 per million output tokens, cheaper than 3.5 Flash's $9.00 output rate.

Q: When is Gemini 4 coming out?
A: Google says Gemini 4 is currently in pre-training, calling it their most ambitious pre-training run yet, with no confirmed release date.

Q: Can developers use it today?
A: Yes, it's available on Google AI Studio and Vertex AI, with limited early access on some consumer interfaces.

好不好用,試了才知道 / Only real use tells you if it's worth it.

Sources / 資料來源

常見問題 FAQ

Gemini 3.6 Flash 是什麼?

Google 於 2026 年 7 月 21 日推出的中階多模態模型,主打更快、更省 token 的 agentic 與程式碼任務表現。

Gemini 3.6 Flash 比 3.5 Flash 好在哪?

輸出 token 減少 17%,DeepSWE 程式碼基準達 49%(3.5 為 37%),OSWorld 電腦操作基準達 83%。

Gemini 3.6 Flash 多少錢?

輸入每百萬 token 1.5 美元,輸出 7.5 美元,比 3.5 Flash 的輸出價格 9 美元更便宜。

Gemini 4 什麼時候出?

Google 表示 Gemini 4 目前在預訓練階段,是他們「最具野心的一次預訓練」,暫無明確發布時間。

一般開發者現在能用嗎?

可在 Google AI Studio、Vertex AI 使用,部分消費端介面也已開放早期存取。

延伸閱讀 / Related Articles


AI 工具觀察站 — 每日精選 AI Agent 與工具趨勢
AI Tool Observer — Daily curated AI Agent & tool trends

留言

這個網誌中的熱門文章

Cursor vs GitHub Copilot vs Claude Code:AI 程式助手大比拼 | AI Coding Assistants Compared: Cursor vs GitHub Copilot vs Claude Code

Stanford 研究登上《Science》:11 個 AI 模型有 47% 機率說你對,即使你錯了 | Stanford Study in Science: AI Models Validate Harmful Behavior 47% of the Time — Sycophancy Is a Real Problem

Claude Code 實測:AI 幫你寫程式到底行不行? | Claude Code Review: Can AI Really Code for You?