Gemini 3.6 Flash 解析:省 17% Token、悄悄預告 Gemini 4 | Gemini 3.6 Flash: Google Cuts Tokens 17%, Teases Gemini 4
By Kit 小克 | AI Tool Observer | 2026-07-23
🇹🇼 Gemini 3.6 Flash 解析:省 17% Token、悄悄預告 Gemini 4
Gemini 3.6 Flash 是 Google 在 2026 年 7 月 21 日悄悄上線的新款中階模型,同時間還推出了 Gemini 3.5 Flash-Lite 與鎖定資安場景的 Gemini 3.5 Flash Cyber。這次更新沒有大型發表會,卻帶來扎實的效能提升:輸出 token 省下 17%、程式碼與電腦操作能力雙雙進步,價格還比上一代更便宜。更值得注意的是,Google 同步透露下一代 Gemini 4 已經進入預訓練階段。
Gemini 3.6 Flash 是什麼?
Gemini 3.6 Flash 是 Google Gemini 系列中主打「速度與推理平衡」的中階模型,鎖定 agentic(自主代理)任務與多模態工作流程,知識截止日為 2026 年 3 月。
跟 Gemini 3.5 Flash 差在哪?
- 輸出效率:平均省下 17% 輸出 token,部分基準測試(如 DeepSWE)甚至省下 65%。
- 程式碼能力:DeepSWE 基準從 37% 提升到 49%。
- 電腦操作:OSWorld-Verified 基準從 78.4% 提升到 83%。
- MLE Bench:從 49.7% 提升到 63.9%。
Gemini 3.6 Flash 定價多少?
輸入每百萬 token 收費 1.5 美元、輸出每百萬 token 收費 7.5 美元,比上一代 Gemini 3.5 Flash 的 9 美元輸出價格便宜不少,對大量呼叫 API 的團隊來說是實質降價。
Gemini 3.5 Flash-Lite 跟 Flash Cyber 又是什麼?
Flash-Lite 主打高吞吐、低延遲,每秒可輸出 350 token,價格只要輸入 0.3 美元、輸出 2.5 美元,部分基準甚至超越上一代旗艦 Gemini 3 Flash(SWE-Bench Pro 54.2% 對 49.6%)。Flash Cyber 則是專門微調來抓資安漏洞的模型,目前只透過 Google 的 CodeMender 代理,開放給政府與合作夥伴做有限試點,反映出 Google 在資安領域的 AI 布局。
Gemini 4 有消息了嗎?
Google 在公告中提到 Gemini 4 目前正在預訓練階段,形容這是他們「最具野心的一次預訓練」,但沒有給出明確發布時間,也沒有透露規模或架構細節。
小克實測心得
Gemini 3.6 Flash 的定位很清楚:不是要打敗旗艦模型,而是把「夠用的智慧」做得更快、更便宜。如果你的應用是大量呼叫 API 的 agentic workflow 或程式碼輔助工具,這次的 token 效率提升是實打實的成本節省,值得優先評估升級。至於 Gemini 4,現在談還太早,但「預訓練已啟動」這句話本身就是個訊號——Google 沒有要在這場模型軍備競賽中放慢腳步。
常見問題 FAQ
Q: Gemini 3.6 Flash 是什麼?
A: Google 於 2026 年 7 月 21 日推出的中階多模態模型,主打更快、更省 token 的 agentic 與程式碼任務表現。
Q: Gemini 3.6 Flash 比 3.5 Flash 好在哪?
A: 輸出 token 減少 17%,DeepSWE 程式碼基準達 49%(3.5 為 37%),OSWorld 電腦操作基準達 83%。
Q: Gemini 3.6 Flash 多少錢?
A: 輸入每百萬 token 1.5 美元,輸出 7.5 美元,比 3.5 Flash 的輸出價格 9 美元更便宜。
Q: Gemini 4 什麼時候出?
A: Google 表示 Gemini 4 目前在預訓練階段,是他們「最具野心的一次預訓練」,暫無明確發布時間。
Q: 一般開發者現在能用嗎?
A: 可在 Google AI Studio、Vertex AI 使用,部分消費端介面也已開放早期存取。
好不好用,試了才知道。
🇺🇸 Gemini 3.6 Flash: Google Cuts Tokens 17%, Teases Gemini 4
Gemini 3.6 Flash is Google's newest mid-tier model, quietly launched on July 21, 2026 alongside Gemini 3.5 Flash-Lite and a cybersecurity-focused Gemini 3.5 Flash Cyber. There was no big keynote, but the upgrade delivers real gains: 17% fewer output tokens, better coding and computer-use scores, and a lower price than its predecessor. Google also quietly confirmed that the next flagship, Gemini 4, is already in pre-training.
What Is Gemini 3.6 Flash?
Gemini 3.6 Flash is Google's mid-tier model built to balance speed and reasoning for agentic tasks and multimodal workflows, with a knowledge cutoff of March 2026.
How Does It Compare to Gemini 3.5 Flash?
- Token efficiency: 17% fewer output tokens on average, up to 65% less on benchmarks like DeepSWE.
- Coding: DeepSWE score jumps from 37% to 49%.
- Computer use: OSWorld-Verified rises from 78.4% to 83%.
- MLE Bench: Improves from 49.7% to 63.9%.
How Much Does Gemini 3.6 Flash Cost?
Pricing is $1.50 per million input tokens and $7.50 per million output tokens — a real price cut from Gemini 3.5 Flash's $9.00 output rate, which matters if you're calling the API at scale.
What About Flash-Lite and Flash Cyber?
Flash-Lite targets high-throughput, low-latency use cases at 350 output tokens/second, priced at $0.30 input / $2.50 output — cheap enough that it beats the older flagship Gemini 3 Flash on some benchmarks (54.2% vs. 49.6% on SWE-Bench Pro). Flash Cyber is a specialized model fine-tuned to find security vulnerabilities, currently deployed only through Google's CodeMender agent in a limited-access pilot for governments and trusted partners.
Is There News on Gemini 4?
Google confirmed Gemini 4 is currently in pre-training, calling it their "most ambitious pre-training run yet" — no release date, size, or architecture details were shared.
Kit's Take
Gemini 3.6 Flash isn't trying to beat the flagship models — it's making "good enough intelligence" faster and cheaper. If you're running agentic workflows or code-assist tools that hammer the API, the token efficiency gains translate into real cost savings worth evaluating now. As for Gemini 4, it's too early to speculate, but "pre-training has started" is itself a signal — Google isn't slowing down in this model arms race.
FAQ
Q: What is Gemini 3.6 Flash?
A: A mid-tier multimodal model Google launched on July 21, 2026, built to be faster and more token-efficient for agentic and coding tasks.
Q: How does Gemini 3.6 Flash improve on 3.5 Flash?
A: 17% fewer output tokens, a DeepSWE coding score of 49% (up from 37%), and an OSWorld computer-use score of 83%.
Q: How much does Gemini 3.6 Flash cost?
A: $1.50 per million input tokens and $7.50 per million output tokens, cheaper than 3.5 Flash's $9.00 output rate.
Q: When is Gemini 4 coming out?
A: Google says Gemini 4 is currently in pre-training, calling it their most ambitious pre-training run yet, with no confirmed release date.
Q: Can developers use it today?
A: Yes, it's available on Google AI Studio and Vertex AI, with limited early access on some consumer interfaces.
好不好用,試了才知道 / Only real use tells you if it's worth it.
Sources / 資料來源
- Google Blog: Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
- 9to5Google: Google launches Gemini 3.6 Flash and 3.5 Flash-Lite, teases Gemini 4
- Artificial Analysis: Gemini 3.6 Flash Intelligence, Performance & Price Analysis
常見問題 FAQ
Gemini 3.6 Flash 是什麼?
Google 於 2026 年 7 月 21 日推出的中階多模態模型,主打更快、更省 token 的 agentic 與程式碼任務表現。
Gemini 3.6 Flash 比 3.5 Flash 好在哪?
輸出 token 減少 17%,DeepSWE 程式碼基準達 49%(3.5 為 37%),OSWorld 電腦操作基準達 83%。
Gemini 3.6 Flash 多少錢?
輸入每百萬 token 1.5 美元,輸出 7.5 美元,比 3.5 Flash 的輸出價格 9 美元更便宜。
Gemini 4 什麼時候出?
Google 表示 Gemini 4 目前在預訓練階段,是他們「最具野心的一次預訓練」,暫無明確發布時間。
一般開發者現在能用嗎?
可在 Google AI Studio、Vertex AI 使用,部分消費端介面也已開放早期存取。
延伸閱讀 / Related Articles
- OpenAI 代理逃出沙箱駭入 Hugging Face:AI 資安事件解讀 | OpenAI Agent Escaped Its Sandbox and Hacked Hugging Face
- CodeMender 解析:Google Gemini AI 資安代理自動修漏洞全解讀 | CodeMender: Google's AI That Auto-Patches Security Bugs
- GPT-5.6 破解30年數學難題:AI真的會做研究了嗎? | GPT-5.6 Solves 30-Year Convex Optimization Problem
AI 工具觀察站 — 每日精選 AI Agent 與工具趨勢
AI Tool Observer — Daily curated AI Agent & tool trends
留言
張貼留言