「3.6 Flash 上线那个晚上」!Google 真正想发的反而没发

7 月 21 日晚上,Logan Kilpatrick 在 X 上发了一条 Vertex AI 更新。r/Bard 用户 MurkhManusya 截了图——截图里那条 Logan 推文写的是:

“Gemini 3.6 Flash is now available on Vertex AI but things are so bleak Logan doesn’t even say ‘Gemini’ anymore”

截图上那条 Vertex AI 推文带 1,591 个 views。MurkhManusya 的 r/Bard 帖子标题起的是 “Lol they doing anything rather than releasing 3.5 pro”,帖子本身 158 票。

但这只是一条截图的事实,不是完整的事实。 我亲自去 Logan X 主页(@OfficialLoganK)抓了一次,发现同时间——7 月 22 日凌晨——Logan 连发了 5 条推文:

  1. “向 Gemini 3.6 Flash 問好” — 103 万 views
  2. “我们对 Gemini 3.5 Flash-Lite 感到非常兴奋” — 19 万 views
  3. “3.5 Flash-Lite 以近乎每秒 350 个输出 token 的速度运行” — 1.9 万 views
  4. “我们已经启动了迄今为止最雄心勃勃的预训练运行,用于 Gemini 4” — 136 万 views
  5. “这个模型效率高得多,并且消耗更少的 token” — 5.8 万 views

也就是说:MurkhManusya 看到的那条 Vertex AI 推文(1,591 views)只是 Logan 同一时间 5 条推文里最小的那条。同时间 Logan 在 Gemini API 渠道发了 4 条明确带 “Gemini” 字的高调推文。Reddit silently 标题只看到了 Vertex AI 那条,没看到 Gemini API 那组。

这是我写这篇的原因——为什么 r/Bard 那 158 票帖子能让“silently”成为标题党,但 Logan 实际上不是 silently。

四个时间点碰在一起

我把过去几个月关于 Gemini 的事件铺开看,发现 4 个时间点撞在一起:

  1. 5 月 19 日(Google I/O 2026):Google 发了 Gemini 3.5 Flash。Ars Technica 当时的描述是“star of the show”。同时承诺 3.5 Pro 会在 6 月发布。
  2. 7 月 16 日(9to5Google):3.5 Pro 仍然没出来。Abner Li 那篇报道的核心信源是 Bloomberg(不是 Business Insider),Bloomberg 引用匿名 Google 内部说“3.5 Pro 在 6 月底重训数据后编码结果 disappointing”。Google 给 9to5Google 的官方回应是“currently testing 3.5 Pro, an upgraded Flash model, and other models with partners”——这句话当时就把“upgraded Flash model”埋进了伏笔。
  3. 7 月 21 日(blog.google 官方公告):Google 跳过 3.5 Pro,直接发 3.6 Flash(+ 3.5 Flash-Lite + 3.5 Flash Cyber)。同一天告诉媒体,3.5 Pro 仍在测试、Gemini 4 预训练已启动。
  4. 7 月 21 日凌晨(blog.google 公告当天):Reddit 用户 MurkhManusya 看到 Vertex AI 那条推文,发了 r/Bard 帖子——“silently”和“doing anything rather than”两个标题都是这天出来的。

把 4 个时间点拼到一起,版本号跳跃是这一连串动作里最反常的一步——Google 没用 “3.5 Flash v2” 或 “3.5 Flash+” 这种补丁式命名,而是直接跳到了 3.6。Ars Technica 当时就指出:3.5 Flash 在 I/O 是“star of the show”,但 2 个月后就被弃用。

注:Logan Kilpatrick 那条 Vertex AI 推文原文我没拉到,本节引用的截图来自 r/Bard 帖子 Lol they doing anything rather than releasing 3.5 pro(MurkhManusya 在 7 月 21 日晚发的,帖子本身 158 票)。截图里的 “1,591 Views” 是 Logan 那条推文当时的浏览数——我没有独立验证这条推文原文,silently 这个论点完全建立在二手截图上。

3.6 Flash vs 3.5 Flash:能力几乎一样

我把 Google 官方 blog 和 Artificial Analysis 两个独立源拉到一起对比。这是核心数据:

指标 3.6 Flash 3.5 Flash 差距
Artificial Analysis Intelligence Index 50 50 0(精确平手)
Output Speed (tokens/s) 275.5 175.7 +57%
Input 价格 ($/1M tokens) $1.50 $1.50 0
Output 价格 ($/1M tokens) $7.50 $9.00 -16.7%
Cache Hit ($/1M tokens) $0.15 $0.15 0
Context window 1M tokens 1M 0
AA Index 评测总成本(一次完整跑) $726.70 $1,040.88 -30.2%
Verbosity(AI Index 跑的 output tokens) 59M 75M -21.3%
DeepSWE 编码(官方 blog 引用) 49% 37% +12 pp
OSWorld-Verified (computer use) 83.0% 78.4% +4.6 pp
MLE-Bench 63.9% 49.7% +14.2 pp
GDPval-AA v2(知识工作) 1421 1349 +5.3%

数据来源:Artificial Analysis Gemini 3.6 Flash 评测页Artificial Analysis Gemini 3.5 Flash 评测页(两表 2026 年 7 月数据)、Google blog 官方公告(Tulsee Doshi 署名,2026-07-21)、Ars Technica 报道(Ryan Whitwam,2026-07-21)、9to5Google 7-16 报道(Abner Li,2026-07-16 12:37 PT,信源是 Bloomberg,不是 Business Insider)。

读完这张表我意识到一件事:

Intelligence 50 vs 50——精确平手,不是“差不多”。 Reddit 主帖标题 “scores the same” 不是夸张——Artificial Analysis 自己给的两个分就是 50 / 50。

Google 几乎所有升级都在“speed / token 效率 / computer use”上:speed +57%(275.5 vs 175.7 tokens/s)、verbosity -21.3%(59M vs 75M tokens)、OSWorld +4.6 pp、DeepSWE 编码 +12 pp、MLE-Bench +14.2 pp。这些是真实升级——但全在执行效率面,不在能力面。

Ars Technica 给的评价是 “marginally more capable”——直接翻译过来就是“稍微更厉害一点”。这个评价跟 AA Intelligence Index 50/50 平手是一致的:能力维度 marginal,效率维度显著。

也就是说:3.6 Flash 的“6”这个版本号,不是技术分水岭,是营销动作。

价格这一栏才是真故事

单看 Output 价格 $7.50 vs $9.00,感觉 Google 在降价。但 r/singularity 上有位 Top 1% Commenter 用户 elemental-mind(专门在 AI 模型定价类帖子下做长分析的评论者)直接揭穿:

“Sorely needed. That 3.5 price hike was overboard. Going from $0.50/3.00 to $1.00/9.00 was totally misplaced. Now they are trying to hide their costs in another $0.50 price hike in input while taking less in output. The selling point of flash is gone.”

翻译:3.5 Flash 早期价格是 input $0.50 / output $3.00——也就是用户实际花的“主力价格”。后来被 Google 涨到 input $1.00 / output $9.00,再到现在 3.6 Flash 的 input $1.50 / output $7.50。

把这条价格史画成表:

时点 Input ($/1M) Output ($/1M) 备注
3.5 Flash 早期 $0.50 $3.00 用户主力价
3.5 Flash 涨价后 $1.00 $9.00 Reddit 公认“overboard”
3.6 Flash(现在) $1.50 $7.50 比 3.5 Flash 末态贵 / 便宜?看场景

elemental-mind 的核心指控:input 从 $1.00 涨到 $1.50(+50%),output 从 $9.00 降到 $7.50(-16.7%),但实际工作里 input tokens 通常比 output 多——这是把“涨价”藏在 input、把“降价”展示在 output 的对称式操纵

但反过来看,r/Bard 用户 Truantee(普通用户,刚来 r/Bard 几周)在 7 月 21 日晚的原话是:

“surprisingly it is cheaper than gemini 3.5 flash!”

为什么两个人都说“便宜”,但角度相反?因为 3.5 Flash 涨价后那个 $1.00/$9.00 是新常态——它跟“$0.50/$3.00 的旧常态“没法直接比。

注:elemental-mind 的早期价格 $0.50/$3.00 我只在 Reddit 评论里见过,没拉到 Google 官方公告里这条定价历史。可能也是 Reddit 用户基于记忆的描述,不一定精确。3.5 Flash 涨价到 $1.00/$9.00 这条是 Google 自己公开过的。

Reddit 的“看法”信号

我把 Reddit / 媒体 / 官方这几类里的“看法”集中放在这里。Reddit 是这次最有信号的——因为它不写新闻稿,它写情绪。

r/singularity 主帖标题的脉络

按从主到次的脉络:

  1. “Gemini 3.6 Flash benchmarks” — 抓 Artificial Analysis 数据的首发帖,本篇 §3 的 AA 数据就是从这里起源
  2. “Google silently released Gemini 3.6 Flash” — Reddit silently 标题的来源——下面 §“渠道分层”会专门讲
  3. “Gemini 3.6 Flash scores the same on Artificial Analysis as 3.5 Flash” — 直接质疑——本篇 §3 也引用了这个观察
  4. “Gemini 3.6 Flash is the fastest frontier model available… by a lot!” — 正面声音——速度领先这个观察也进了 §3

r/Bard 主帖标题的情绪梯度

  1. “Gemini 3.6 Flash released On AI Studio” — 首发帖,事实型
  2. “Lol they doing anything rather than releasing 3.5 pro” — 标题即是情绪,158 票,r/Bard 帖子整体情绪代表
  3. “Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber” — 转官方 blog,事实型
  4. “3.6 now has ‘Unknown’ knowledge cutoff. Certified DeepMind moment” — 知识截止日期缺失被截图吐槽
  5. “I asked gemini 3.6 flash which model is best for coding” — 问模型对自身评分的“诱导测试”,具体结果我没读
  6. “Gemini 3.5 flash-lite and 3.6 flash compare in artificial analysis” — Lite vs Flash 的具体对比

剩下几条 r/Bard 帖子(“Gemini 3.6 flash benchmark is out!” 等)信息密度低,不展开。

r/LocalLLaMA 0 命中

这个信号可能比上面所有加起来都重要:r/LocalLLaMA 本周 top 帖是 Thinking Machines Inkling(开源权重)、Kimi K3、Kimi-K3 vs Fable、Qwen 3.6、Macaron-V1-Venti——一个 Gemini 相关帖子都没有。

r/LocalLLaMA 是关注“开源大模型”的社区——他们本来就不在意闭源模型。但当一个 7.4 万 stars 量级的 Paperclip AI 都在本周持续出现时,Google 自家 Gemini 3.6 Flash 在闭源圈外的存在感是显著低于 Inkling 这种开源新模型。

r/MachineLearning 0 命中

学术界也没讨论。一个商用闭源模型的发布在 r/MachineLearning 上完全没激起水花,跟 r/singularity 形成对比——singularity 是关注 AGI 时间表的社区,他们会跟进;但 r/ML 是技术深度的社区,他们不在乎。

“渠道分层”才是这件事真正值得记住的词

回到 Logan 那 5 条推文。

我前面贴了:同时间 Logan 在 X 上发了 5 条——4 条带 “Gemini” 字(向 Gemini 3.6 Flash 問好 / Gemini 3.5 Flash-Lite / Gemini 4 / token 效率),1 条不带(Vertex AI 公告)。

Reddit 用户 MurkhManusya 只看到 Vertex AI 那条——这是截图的局限,不是 Logan 行为的全部。Logan 不是 silently,而是在两个渠道用两种语气发同一件事

  • Vertex AI 渠道(面向企业客户):措辞保守,“is now available” + 三个参数——更像运维公告
  • Gemini API 渠道(面向开发者):满 Gemini 字 + “对 Gemini 3.5 Flash-Lite 感到非常兴奋” + 350 tokens/s 性能 — 营销级语气

这是 Google 把企业客户和开发者用不同语气运营——这不是 silently,是渠道分层。Reddit 那个 158 票 “silently” 标题是观察者偏见:用户从一张截图反推“Google 不想提 Gemini”,但实际 Logan 在另一个渠道最响亮地提了 Gemini。

但 r/singularity 用户还是从这条截图里推出有意义的争论——这才是 Reddit 真正的价值。Hereitisguys9888(在 r/singularity 长期跟踪 Google AI 发布节奏的观察者,7 月 22 日凌晨)的原话:

“At this point, gemini 3.5 pro has to be a breakthrough. Even if it’s on fable level, its way too late considering open source is reaching fable level”

duluoz1(同样是 Google AI 发布跟踪者,7 月 22 日凌晨)紧跟一句:

“Yup. This is why they’re delaying. If they release their new model and it’s behind the open weight latest models, they’ll be impact to their share price”

这两条加在一起是个完整论证:Google 不发 3.5 Pro,不是因为它没做完,是因为做完发出来如果打不过 Fable 5 / Inkling 这种前沿模型,对 Alphabet 股价有直接影响。——这个论证跟 silently 无关,而是基于 Bloomberg 报道的编码延迟事实。

但这个论点有个破绽:FuzzyBucks(在 r/singularity 有声望的长期评论者,专门盯 Google 战略动向,7 月 22 日凌晨)反驳:

“Imo, they admitted ‘defeat’ earlier this year when they basically stopped serving and then shut down Gemini CLI and CodeAssist. Their investment in Anthropic is how they are staying in the ‘big leagues’. They bought 10% of anthropic at the same time as they stopped serving the tools I mentioned. Altogether, they own ~14% of Anthropic”

按这条说法,Google 已经“弃疗”——Gemini CLI 和 CodeAssist 关掉,10% + 既有 = 共持 Anthropic ~14% 股权。

Greedyanda(另一位 r/singularity 长期评论者,更偏产品技术观察,7 月 22 日凌晨)又反驳 FuzzyBucks:

“They shut down Gemini CLI because there is now Antigravity CLI.”

两个反驳都基于“读 Logan 行为”的解读,但走向相反:FuzzyBucks 说“Google 在战略撤退”,Greedyanda 说“Google 在做产品替换”。 现在再加上我前面发现的 Logan 5 条推文:Logan 在 Vertex AI 渠道保守、在 Gemini API 渠道高调——这更接近 Greedyanda 的“产品替换”解读,而不是 FuzzyBucks 的“战略撤退”。

那 Google 到底在想什么

把上面所有信号放一起,我能拼出三种解释——它们不一定互斥,但优先级不同:

第一种:FuzzyBucks 的“战略撤退说”

证据链:

  • Gemini CLI 关闭
  • CodeAssist 关闭
  • 3.5 Pro 持续延迟
  • Google 持 Anthropic ~14%
  • Antigravity CLI 上线(但 r/singularity 用户 nemzylannister(开发者,长期在 r/singularity 发编码 / 工具类评论)直接说 “theres no way you use antigravity. wtf do you do with it? it cant do a single thing for me”——原评论含粗口,已清洗)

反证:Google 同时在大力投入 Isomorphic Labs(Demis Hassabis 联合创办的药物 AI)、AlphaFold 系列、world models、robotics。LLM 不是 Google 唯一押注。

第二种:Hereitisguys9888 的“打不过开源”说

证据链:

  • 3.5 Pro 编码延迟(Bloomberg 独家报道 + Google 官方“currently testing”回应)
  • Inkling / Kimi K3 / Qwen 3.6 本周在 r/LocalLLaMA 占据前列
  • Artificial Analysis 50/50 精确平手说明 3.6 Flash 没把能力分推高

反证:如果 Google 真打不过开源前沿,那 3.6 Flash 在 DeepSWE +12 pp 就是 AI 巨头仍在追赶的事实。慢,不是停。

第三种:virualQubit(r/Bard 7 月 22 日上午)的“长线说”——一位把“Google = AGI 长跑者”叙事认真执行的人

“Nah, don’t underestimate them. Maybe it’s a difficult moment, but they’re way too big and powerful to give up like that. They are the pioneer of AI, don’t forget the iconic paper ‘Attention is all you need’ by Google. They are working on world models, video models, music models, robotics, they are playing the long game.”

证据链:

  • Gemini 4 预训练已启动(官方 blog 自己说)
  • AlphaFold 系列 / Isomorphic Labs / 世界模型
  • “Attention is all you need” 是 Google 2017 年自己发的

反证:长线是真的,但 Alphabet 投资者要的是季度财报,不是 5 年后的机器人。“长线”不等于“这次不发 Pro 的解释”。

我自己的判断是 A+B 的混合

  • Google 在做长线(C),但
  • 短期面对的是“打不过开源 / Anthropic 前沿”的尴尬(A+B),所以
  • 战略上把“前沿 LLM”押给 Anthropic 联盟 + 自己的 Antigravity / Workspace 集成(这就是 FuzzyBucks 看到的“弃疗”信号)
  • 同时把“性价比 + 速度 + 生态绑定”作为 3.6 Flash 的牌

3.6 Flash 不是“模型升级”,是“战略转向的产品化”。

我没去验证的事

为了避免变成单方面叙事,我承认:

  • Logan Kilpatrick X 原帖已拉——他 7 月 22 日凌晨在 @OfficialLoganK 上连发 5 条推文:4 条带 “Gemini” 字(向 Gemini 3.6 Flash 問好 103 万 views / Gemini 3.5 Flash-Lite 19 万 / Gemini 4 预训练 136 万 / token 效率 5.8 万),1 条 Vertex AI 公告(1,591 views)。Reddit 用户 MurkhManusya 截的图只显示了 Vertex AI 那条,就推导出 “silently”——本篇 §6 已据此重写为“渠道分层”观察。
  • “3.5 Flash 早期 $0.50/$3.00 价格”只在 Reddit 评论里见过——没在 Google 官方公告里找到这条定价历史记录。可能 Reddit 用户基于记忆的描述,不一定精确。
  • “Google 持 Anthropic ~14%“是 r/singularity 用户 FuzzyBucks 的数字——没在 Alphabet 公开财报里独立验证。Google 之前公告过投资 Anthropic $2B(2023 年)+ 后续加码,但 14% 这个具体比例需要查最新 13F 或 Anthropic 估值。
  • Business Insider / Bloomberg 3.5 Pro 跳票独家是付费墙——我没直接读到 Bloomberg 原文,只看到 9to5Google (Abner Li) 的转述 + Ars Technica 引用的“earlier reports claimed”措辞。9to5Google 那条 7-16 报道已经确认是 Bloomberg 信源,不是 Business Insider。
  • Antigravity CLI / CodeMender 实际能力我没亲测——只有 Reddit 用户吐槽和官方 demo 视频。
  • Artificial Analysis 3.5 Flash 的精确分值:Intelligence 50 / Speed 175.7 tokens/s / Verbosity 75M tokens / 评测成本 $1,040.88,Intelligence Index 跟 3.6 Flash 精确平手(50 vs 50)。Reddit 用户 Mysterious_Bed_1804 之前说 “3.6 scores just below 3.5”——该 Reddit 评论是错的,AA 数据是 50/50 平手。
  • Greg Isenberg 47 分钟访谈 transcript 里 0 命中 ‘autonomous economy’ ‘GDP’ ‘country’ 这件事——是上一篇文章查 Paperclip 时的发现,本文不展开但保留为边角数据点。
  • 本篇没拉 r/singularity 主帖评论里的中文 / 西语 / 葡语用户——MetalZone00 (西语) 说“No puede hacer nada contra los chinos”,GoldPossession7284 (葡语) 说 3.5 Pro 在测试 + Gemini 4 已开始预训练——这些信号我只在标题层面看到,没深入读全文评论。

我接下来想做的

  • 等 3.5 Pro——如果 8 月发出来,看是否真的“对标 GPT 5.6 和 Claude Fable/Sonnet 5”(Ars Technica 原话)
  • 等 Gemini 4 任何信号——预训练已启动,但什么时候出 V1 没有任何时间表
  • 监控 Logan Kilpatrick 接下来 30 天的 X——如果他继续 Vertex AI / Gemini API 双渠道分发,渠道分层就是长期策略
  • 把 r/singularity “silently” 主帖评论读完——目前只读了 top 10,可能漏了高赞长文
  • 跑一次 3.6 Flash API 实测,看 Reddit 用户 ozone6587 的“dumber models useful for coding or agentic tasks”判断在自己用例里成不成立

PS:写这篇的时候我一直在想一件事——Google 用 3.5 → 3.6 这一步,把“Flash 系列升级”重新包装成“代际跳跃”,同时把“3.5 Pro 没发”这件事藏在版本号跳跃的噪声里。这不是新招,用版本号转移注意力是商业叙事的常用工具——关键是读者别被版本号牵着走,要看能力 / 价格 / 信号本身。

PPS:本篇所有 Reddit 评论引用都来自 r/singularity 主帖 Google silently released Gemini 3.6 Flash(293 票)和 r/Bard 主帖 Lol they doing anything rather than releasing 3.5 pro(158 票)的 top 10 评论。我没读完整 80-123 条评论,只读了 top 10。低赞评论里可能还有更尖锐的反向声音。


数据来源(本稿数据抓取时点 2026-07-22 下午,所有 Reddit / 媒体帖的“小时前”都是相对这个时点):


相关阅读《「数据没铺开就是幻觉」》 — 同一套“全量铺开再归纳”的方法论起点:《「一个真正在跑的 AI 公司长什么样」》 — 同一套“7 个独立源”做法套到 Paperclip 上的产物。

💬 留言

留言使用 Giscus, 通过 GitHub Discussions 驱动。需 GitHub 账号,无需注册。