Signal Brief

Qwen-Image-3.0 发布

Alibaba Qwen 团队发布了第三代图像生成基座模型 Qwen-Image-3.0。该模型支持高达 4.5k tokens 的长提示词,可实现复杂的布局生成(如报纸、分镜脚本、3x3 信息图),具备 10px 级别清晰的文本渲染能力,并支持 12 种语言和 100 多种艺术风格。

twitter关注列表 Qwen (@Alibaba_Qwen) 发布 2026-07-22 收录 2026-07-22 观察

一句话判断

重点关注其对复杂布局、长文本理解及嵌套 UI 渲染的生产力突破。

核心信息

Alibaba Qwen 团队发布了第三代图像生成基座模型 Qwen-Image-3.0。该模型支持高达 4.5k tokens 的长提示词,可实现复杂的布局生成(如报纸、分镜脚本、3x3 信息图),具备 10px 级别清晰的文本渲染能力,并支持 12 种语言和 100 多种艺术风格。

原始内容

🎨 Meet Qwen-Image-3.0 — the third generation of our foundational image generation model. If 1.0 was about "Precision," and 2.0 added "Variety, Completeness, Beauty & Authenticity," then 3.0 comes down to a single word: Real (实). Three dimensions of "Real": 📰 Rich Content — prompts up to 4.5k tokens. One-pass generation of complex layouts: newspapers, storyboards, exam papers — even a 3×3 infographic grid or picture-in-picture-in-picture UIs. 🔬 Authentic Details — text legible down to 10px, full LaTeX paper pages, pores, hair strands & near-photographic skin texture. 🌏 Deep Knowledge — native rendering in 12 languages, 100+ art styles, realistic UIs (web / games / livestreams), plus world knowledge & live web retrieval. Not just "good-looking" — genuinely useful. Image generation as a real productivity tool for design, content, education & e-commerce. Go create 🏃🎨 💬Qwen Chat: https://t.co/941HmITJ2W 📝Blog: https://t.co/5mnS4uI9Ar ![photo](https://pbs.twimg.com/media/HN1LmIqa8AATJuu.jpg) Qwen (@Alibaba_Qwen): Authentic Details and Deep Knowledge are also major upgrades in this release. Together, they enable Qwen-Image-3.0 to make real breakthroughs in high-value productivity scenarios — newspaper PDFs, short-drama storyboards, complex UI, and more. As image generation keeps advancing, we believe it will unlock genuine productivity value across design, content creation, education, e-commerce, and beyond. Qwen (@Alibaba_Qwen): Beyond horizontal expansion, depth is another key trait of Rich Content — testing the model's semantic deconstruction and logical nesting: rendering multiple nested interfaces layer by layer within a single image. The example below uses one instruction to display, from outer to inner: VSCode → Qwen Chat → a messaging app → a pour-over coffee poster. Each layer preserves the authentic style of its UI, forming a "picture-in-picture-in-picture" visual depth.

相关动态

01

百度开源 Unlimited-OCR 再登 HuggingFace 前三

百度开源的无限制OCR模型Unlimited-OCR(30亿参数,32K上下文)再次登上HuggingFace总榜第三,被Yann LeCun转发。该模型可一次性解析100页PDF,40页后错误率低于0.11,准确率93%,GitHub Star 1.65万,HuggingFace下载量224万。技术核心是R-SWA机制,能保持恒定KV ...

twitter关注列表2026-07-22#模型发布#开源#技术突破
观察
02

Cosmos 3 Super模型发布

NVIDIA AI 发布了 4 步 Cosmos 3 Super 模型,生成图像和视频速度比原版快 25 倍,在 Artificial Analysis 基准中图像到视频(无音频)排名第一,文本到图像排名第二,模型已在 Hugging Face 开源。

twitter关注列表2026-07-22#模型发布#技术更新#AI
观察
03

Codex 1000万付费用户推动OpenAI进军企业市场

OpenAI的Codex已突破1000万付费用户,其中大量用户选择Pro层套餐,这直接冲击了Anthropic在B2B和企业市场的收入优势。文章指出,OpenAI正通过Codex向高价值专业客户市场渗透,而Anthropic因竞争压力被迫保留Fable 5在Max层套餐中,这种竞争有望加速Opus 5和Fable 5.5的发布。

twitter关注列表2026-07-22#模型发布#行业动态
观察
05

AI解决87年雅各比猜想

Levent Alpöge在Anthropic完成对87年前未解决数学猜想的反例发现,证明yaqūjīzé jiāujiàn(雅各比猜想)存在无法恢复单一输入的多项式映射,验证已通过Lean数学形式化检查。与人类研究者相比,AI通过计算数学探索发现此反例,展现数学证明领域AI能力增强。

twitter关注列表2026-07-22#技术突破#数学#模型发布
值得跟进