Signal Brief

Google 发布 Gemini Omni Flash 登顶视频生成排行榜

Google 发布 Gemini Omni Flash 多模态模型,在 Artificial Analysis 视频排行榜上文本到视频和图像到视频均位居第一,击败 ByteDance 的 Seedance 2.0。模型支持文本、图像、视频输入,生成 720p 24FPS 3-10 秒视频,定价 $0...

twitter关注列表 Google (@Google) 发布 2026-07-16 收录 2026-07-16 观察

一句话判断

该模型在视频生成任务中取得领先排名,且定价与现有方案一致,值得尝试并关注其后续迭代。

核心信息

Google 发布 Gemini Omni Flash 多模态模型,在 Artificial Analysis 视频排行榜上文本到视频和图像到视频均位居第一,击败 ByteDance 的 Seedance 2.0。模型支持文本、图像、视频输入,生成 720p 24FPS 3-10 秒视频,定价 $0.10/秒,已通过 Gemini API、Google AI Studio 等平台开放使用。

原始内容

Google (@Google) 转发了 News from Google (@NewsFromGoogle) 的帖子: Gemini Omni Flash took the #1 spot on the @ArtificialAnlys Leaderboards for Text to Video and Image to Video🏆 Omni Flash is our latest multimodal model that allows you to create and edit high-quality videos just by describing what you want to see. Try it out in the @GeminiApp, @FlowByGoogle and via our APIs. > **引用原帖 Artificial Analysis (@ArtificialAnlys):** > Google's Gemini Omni Flash debuts at #1 on the Artificial Analysis Text to Video and Image to Video Leaderboards, edging out ByteDance's Seedance 2.0 on both > Gemini Omni Flash is the first model in Google's Gemini Omni family, unveiled at Google I/O in May and opened to developers in public preview on June 30. Google positions Omni as a natively multimodal model that can "create anything from any input", starting with video: it accepts text, images, and video as input, generates clips with native audio, and supports conversational editing, where prompts change a video while preserving the rest of the scene. Gemini Omni Flash generates 3 to 10 second clips at 720p and 24 FPS, in 16:9 or 9:16, with longer durations coming soon. > In the Artificial Analysis Video Arena, Gemini Omni Flash debuts at #1 on both the Text to Video and Image to Video Leaderboards, narrowly ahead of ByteDance's Seedance 2.0 on each. > Gemini Omni Flash is priced at $0.10 per second of generated video ($6.00 per minute), matching Veo 3.1 Fast. The rate is the same for Text to Video and Image to Video. > It is available now in the Gemini API, Google AI Studio, and the Gemini Enterprise Agent Platform, in the Gemini app and Google Flow for consumers, and at no cost in YouTube Shorts and the YouTube Create app. > Congratulations to @GoogleDeepMind on the release! > See below for comparisons between Gemini Omni Flash and other leading models in the Artificial Analysis Video Arena 🧵 > https://x.com/ArtificialAnlys/status/2076747075036045645

相关动态

01

Kimi K3 在 SpreadsheetBench 2 上排名第一

Kimi K3 在 SpreadsheetBench 2 基准测试中排名第一,超越 Claude Fable 5,完成 34.8% 的 workflow 任务。该基准测试涉及 321 个专家策划任务,平均每个任务包含 11.8 个工作表和 593.5 个单元格更改,聚焦于完整的电子表格工作簿执行。

twitter关注列表2026-07-18#模型#技术突破#评测
观察
02

BestBlogs 早报 · 07-18

月之暗面发布Kimi K3,2.8万亿参数,896选16的Stable LatentMoE,上下文100万token,接近Fable-5但仍落后最强闭源模型,完整权重7月27日前开源;VentureBeat调查显示54%企业已发生AI代理安全事件;xAI开源Grok Build(84万行Rust代码)并残留上传用户代码痕迹;Cursor评...

twitter关注列表2026-07-17#AI#模型发布#开源
观察
03

Kimi K3 编码代理评测

Artificial Analysis发布Kimi K3在编码代理指数上的评测结果:得分57,排名第5,性能与GPT-5.6 Terra和GPT-5.5持平(57),超过Opus 4.8(55),略低于Grok 4.5(58)和Fable 5(59);成本平均3.18美元/任务,比GPT-5.6 Sol便宜55%。

twitter关注列表2026-07-17#AI模型#评测
观察
05

Runway Agent 在第三方评测中全面领先

Physion Labs 发布 Physion-Arc 1.0 基准测试,对 Runway、Luma、MiniMax、Kling、Utopia 和 TapNow 六款 AI 视频代理进行独立人类评估,使用 30 个电影提示和 16 个指标。Runway Agent 2.0 在叙事连贯性、电影语言和制作质量三项核心维度全部排名第一。

twitter关注列表2026-07-17#评测#多模态#模型
观察