发生了什么
Google DeepMind 发布 Gemini 3.6 Flash 等三款模型(65% token 节省, 350 tokens/sec)
为什么值得关注
与 Gemini 3.5 Flash 相比,3.6 Flash 在相同成本下节省 65% token,且 3.5 Flash-Lite 速度提升至 350 tokens/sec,值得原文确认细节。
信息来源
以下内容来自公开来源,可打开原文继续核验。
Google DeepMind 发布三款新模型
Google DeepMind 发布 Gemini 3.6 Flash、3.5 Flash-Lite 和 3.5 Flash Cyber 三款新模型,其中 3.6 Flash 相比 3.5 Flash 在复杂编程任务中 token 使用减少 65%,3.5 Flash-Lite 输出速度达 350 tokens/sec,两款模型已在 Gemini 应用上线,同时 3.5 Pro 进入合作伙伴测试。
打开原始来源 ↗Gemini 3.6 Flash 主打更高 token 效率
Google DeepMind 推出 Gemini 3.6 Flash 模型,称其在相同成本下使用更少的 token 即可提供更高质量输出。同时发布 3.5 Flash-Lite 用于日常任务和 3.5 Flash Cyber 用于网络安全。
打开原始来源 ↗