Signal Brief

Sonilo 发布音频模型 1.0

Sonilo 发布了视频原生音频模型 Sound Effects 1.0,可自动从视频同步音频,支持三分钟输入并可根据文本生成独立音频素材。该模型与现有的 Video2Music 和 Text2Music 工具集成,Fal 独家提供 Day Zero API,并在基准测试中声称超越领先模型。

twitter关注列表 🚨 AI News | TestingCatalog (@testingcatalog) 发布 2026-07-23 收录 2026-07-23 观察

一句话判断

新增视频原生音频生成和自动同步功能,同时支持文本驱动的独立音频素材生成,并在基准测试中超越领先模型。

核心信息

Sonilo 发布了视频原生音频模型 Sound Effects 1.0,可自动从视频同步音频,支持三分钟输入并可根据文本生成独立音频素材。该模型与现有的 Video2Music 和 Text2Music 工具集成,Fal 独家提供 Day Zero API,并在基准测试中声称超越领先模型。

原始内容

Sonilo has released Sound Effects 1.0, a video-native audio model that can read on-screen motion, scene context, and timing. > It can make sound effects synced to the footage. > Video-to-Sound-Effects runs automatically without requiring a prompt. > Text-to-Sound-Effects is able to generate standalone assets from a written description. https://video.twimg.com/amplify_video/2080415331688775680/vid/avc1/1280x720/gRKpKwiYcF1JepCv.mp4?tag=29 > **引用原帖 Sonilo (@Sonilo_music):** > Meet the first Sound World Model from Sonilo. > Generated video can look incredible, but without the right sound, it still feels distant. > Sonilo takes a video and creates the music and sound effects it needs, timed to the scene, motion, mood, and environment. > And in head-to-head benchmarks, Sonilo outperforms leading models across both music and sound effects. > With Sound Effects 1.0, we’re bringing music and sound effects together into one sound layer for videos. > Music carries emotion. > Sound effects create presence. > Together, they make worlds feel real, immersive, and alive! > Sound Effects 1.0 is live. > #SoundWorldModel #Sonilo #AIAudio > https://x.com/Sonilo_music/status/2080337253046595673 🚨 AI News | TestingCatalog (@testingcatalog): Users can also add a prompt to shape what the model generates. The footage still dictates when each sound is added. Video inputs can run up to three minutes, and the model pairs with Sonilo's Video2Music and Text2Music, meaning the same upload can generate sound effects and music. Day zero API access ships through Fal as the exclusive launch partner. Test it out 👀 https://t.co/H4z28UPgp5

相关动态

02

FLUX 3 将极大影响机器人智能和学习

Black Forest Labs 发布 FLUX 3 统一多模态模型,支持图像、视频、音频和动作预测。其 Self-Flow 架构使机器人微调数据减少一半,相比早期模仿视频实验,机器人训练数据减少 10 倍,通过大规模视频预训练隐式学习物理规律。

twitter关注列表2026-07-23#模型发布#多模态#技术突破
观察
03

Black Forest Labs 发布 FLUX 3 统一多模态模型

Black Forest Labs 发布 FLUX 3,采用统一架构同时处理图像生成、视频、原生音频和机器人动作预测。视频已开放早期访问,后续将开放权重,并与 mimic、Audi 合作在真实机器人上运行。团队认为物理世界智能与内容创作可共享同一视觉基础模型。

twitter关注列表2026-07-23#模型发布#多模态#技术突破
观察