Signal Brief

Gemini 3.5 Flash-Lite 发布

Logan Kilpatrick 宣布 Google 发布 Gemini 3.5 Flash-Lite,这是最小最快的 Gemini 模型,速度近 350 tokens/s,比 Gemini 3 更智能,与 Gemini 2.5 Flash 同成本但更智能,超越 3.1 Flash-Lite。

twitter关注列表 Logan Kilpatrick (@OfficialLoganK) 发布 2026-07-21 收录 2026-07-21 观察

一句话判断

该模型速度达 350 tokens/s,适合对延迟敏感的 UI 和 agent 场景,值得关注其实际应用效果。

核心信息

Logan Kilpatrick 宣布 Google 发布 Gemini 3.5 Flash-Lite,这是最小最快的 Gemini 模型,速度近 350 tokens/s,比 Gemini 3 更智能,与 Gemini 2.5 Flash 同成本但更智能,超越 3.1 Flash-Lite。

原始内容

I am very excited about Gemini 3.5 Flash-Lite, our smallest and fastest Gemini model! - it is more intelligent in many cases than Gemini 3 - same cost and smarter than Gemini 2.5 Flash (which is approaching end of life) - also out paces 3.1 Flash-Lite on most use cases! https://t.co/tJd2tDmyac ![photo](https://pbs.twimg.com/media/HNwp4EdbIAAm2qY.jpg) ![photo](https://pbs.twimg.com/media/HNwp9RFacAAVgMM.jpg) Logan Kilpatrick (@OfficialLoganK): 3.5 Flash-Lite runs at nearly 350 output tokens per second which feels so smooth on many latency sensitive UI experiences and is now also a viable option to drive agent harnesses!

相关动态

02

NVIDIA Vera Rubin 平台发布

NVIDIA 发布 Vera Rubin 平台,性能功耗比提升 10 倍;CoreWeave 等云服务商部署 Vera Rubin NVL72,每兆瓦 token 数比 Blackwell 多 10 倍;DeepInfra 基准测试显示 Vera CPU 速度是其他 CPU 的 2 倍以上,支持更多并发 AI 代理。

twitter关注列表2026-07-21#技术突破#产品发布#AI
值得跟进
05

Grok 4.5 集成 Microsoft Outlook

Grok 4.5 现可直接在 Microsoft Outlook 中使用,支持总结邮件线程与附件、识别决策与待办任务、以用户语气起草回复、搜索网络和 𝕏,以及整理、归档、删除或标记邮件。

twitter关注列表2026-07-21#AI#产品更新#大模型
观察