Signal Brief

Artificial Analysis 评测 Gemini 3.6 Flash 和 3.5 Flash-Lite

Artificial Analysis 发布了对 Google Gemini 3.6 Flash 和 Gemini 3.5 Flash-Lite 的基准测试结果。Gemini 3.6 Flash 智能指数保持 50 未变,但时间减半至 1.3 分钟,成本下降 18% 至 $0.50;Gemini 3...

twitter关注列表 Artificial Analysis (@ArtificialAnlys) 发布 2026-07-21 收录 2026-07-21 观察

一句话判断

该评测提供了详细的基准数据对比,有助于模型选型决策,值得点开原文查看完整分析。

核心信息

Artificial Analysis 发布了对 Google Gemini 3.6 Flash 和 Gemini 3.5 Flash-Lite 的基准测试结果。Gemini 3.6 Flash 智能指数保持 50 未变,但时间减半至 1.3 分钟,成本下降 18% 至 $0.50;Gemini 3.5 Flash-Lite 智能指数提升 11 分至 36,时间减半至 0.6 分钟,但成本翻倍至 $0.09。

原始内容

Google has released Gemini 3.6 Flash and Gemini 3.5 Flash-Lite. Both halve time per task relative to their predecessors and increase token efficiency, Gemini 3.5 Flash-Lite improves by 11 Intelligence Index points while Gemini 3.6 Flash does not improve in intelligence over 3.5 Flash @GoogleDeepMind has released the latest updates to the Gemini model family with two new models. We benchmarked Gemini 3.6 Flash and Gemini 3.5 Flash-Lite ahead of release across Intelligence, Time per Task, and Cost per Task Key takeaways for Gemini 3.6 Flash (high reasoning): ➤ Maintains the same Intelligence as Gemini 3.5 Flash: Gemini 3.6 Flash scores 50 on the Artificial Analysis Intelligence Index, matching Gemini 3.5 Flash, and just below recently released models Muse Spark 1.1 (xhigh, 51) and GPT-5.6 Luna (max, 51). Compared to Gemini 3.5 Flash, Gemini 3.6 Flash maintains similar scores across the Index, with an improvement in GDPval-AA v2 (1421, +72) and a slight regression in HLE (38%, -3 points) ➤ Half the Time per Task: Gemini 3.6 Flash records an average time per task of 1.3 minutes, a more than 50% reduction compared to Gemini 3.5 Flash (2.7). This is driven by increased token efficiency and faster output, with speeds measured at 304 output tokens per second in our pre-launch testing ➤ Slightly lower Cost per Task: Gemini 3.6 Flash’s cost per task decreases ~18%, from $0.59 to $0.50. This is driven by lower output token use and new pricing of $1.50/$7.50 per 1M input/output tokens, down from $1.50/$9.00 for Gemini 3.5 Flash Key takeaways for Gemini 3.5 Flash-Lite (high reasoning): ➤ Significant Intelligence improvements over Gemini 3.1 Flash Lite: Gemini 3.5 Flash-Lite scores 36 on the Artificial Analysis Intelligence Index, up 11 points from Gemini 3.1 Flash-Lite (25). This places it behind models such as Nemotron 3 Ultra (38) and DeepSeek V4 Flash (max, 40), and above Mistral Medium 3.5 (30). The biggest intelligence gains compared to Gemini 3.1 Flash-Lite are in agentic evaluations, with improvements in GDPval-AA v2 (1140, +498), TerminalBench v2.1 (53.6, +22.5 points) and Tau3-Banking (16.5%, +7.8 points) ➤ Nearly half the Time per Task: Gemini 3.5 Flash-Lite records an average time per task of 0.6 minutes, nearly a 50% reduction compared to Gemini 3.1 Flash-Lite (1.0). This is driven by increased token efficiency and fast output speed, measured at 350 output tokens per second in our pre-launch testing ➤ More expensive with 2x Cost per Task: Gemini 3.5 Flash-Lite’s average cost per task increases from $0.04 to $0.09, driven by new pricing of $0.30/$2.50 per 1M input/output tokens, up from $0.25/$1.50 for Gemini 3.1 Flash-Lite. This cost increase comes despite using fewer output tokens, falling from 20k to 13k average output tokens per task Key model details: ➤ Context window: Both models retain the same 1M context window as their predecessors ➤ Multimodality: Both models have text, image, video, and speech input with text output only ➤ Pricing: Gemini 3.6 Flash is priced at $1.50/$7.50 per million input/output tokens, down from Gemini 3.5 Flash at $1.50/$9.00. Gemini 3.5 Flash-Lite is priced at $0.30/$2.50 per million input/output tokens, with the same input pricing across all input modalities. This is an increase from Gemini 3.1 Flash-Lite, which is priced at $0.25/$1.50 per million input/output tokens, with input audio tokens at $0.50. Both models retain the same 90% discount for cached input tokens ![photo](https://pbs.twimg.com/media/HNw0NHUaEAAMfjy.jpg) Artificial Analysis (@ArtificialAnlys): Full benchmark results for Gemini 3.6 Flash (high) and Gemini 3.5 Flash-Lite (high) https://t.co/TuLEOWjzPU Artificial Analysis (@ArtificialAnlys): For further analysis, see https://t.co/xpijJJ9D6f

相关动态

01

NVIDIA Vera Rubin 平台发布

NVIDIA 发布 Vera Rubin 平台,性能功耗比提升 10 倍;CoreWeave 等云服务商部署 Vera Rubin NVL72,每兆瓦 token 数比 Blackwell 多 10 倍;DeepInfra 基准测试显示 Vera CPU 速度是其他 CPU 的 2 倍以上,支持更多并发 AI 代理。

twitter关注列表2026-07-21#技术突破#产品发布#AI
值得跟进
03

Grok 4.5 集成 Microsoft Outlook

Grok 4.5 现可直接在 Microsoft Outlook 中使用,支持总结邮件线程与附件、识别决策与待办任务、以用户语气起草回复、搜索网络和 𝕏,以及整理、归档、删除或标记邮件。

twitter关注列表2026-07-21#AI#产品更新#大模型
观察
04

Anthropic 1.5亿美元和解作者版权诉讼

美国法官批准Anthropic支付1.5亿美元给作者,作为其使用盗版书籍训练Claude的版权和解金。这是美国版权案件中最大和解,涉及700万本盗版书,91%作者已领取份额。法院裁定训练属于合理使用,但存储盗版副本构成侵权。

twitter关注列表2026-07-21#AI#行业动态#政策
观察
05

Gemini 3.5 Flash-Lite 发布

Logan Kilpatrick 宣布 Google 发布 Gemini 3.5 Flash-Lite,这是最小最快的 Gemini 模型,速度近 350 tokens/s,比 Gemini 3 更智能,与 Gemini 2.5 Flash 同成本但更智能,超越 3.1 Flash-Lite。

twitter关注列表2026-07-21#AI#模型发布#大模型
观察