Signal Brief

Grok 4.5成本效率对比评测

Artificial Analysis数据显示,Grok 4.5每次任务成本仅0.31美元,而Claude Fable 5为2.75美元、Claude Opus 4.8为1.80美元、GPT-5.6 Sol为1.04美元、Kimi K3为0.95美元,Grok 4.5成本分别低约9倍、6倍、3倍和3...

twitter关注列表 Elon Musk (@elonmusk) 发布 2026-07-17 收录 2026-07-17 观察

一句话判断

提供了Grok 4.5与竞品在成本和token效率上的量化对比数据,值得阅读原文了解xAI在成本控制上的技术细节。

核心信息

Artificial Analysis数据显示,Grok 4.5每次任务成本仅0.31美元,而Claude Fable 5为2.75美元、Claude Opus 4.8为1.80美元、GPT-5.6 Sol为1.04美元、Kimi K3为0.95美元,Grok 4.5成本分别低约9倍、6倍、3倍和3倍;在FrontierSWE上,Grok 4.5每个任务的总token消耗比Fable 5少6倍,性能仍接近顶端。

原始内容

Grok 4.5 is arguably #1 when taking speed and cost into account > **引用原帖 X Freeze (@XFreeze):** > Grok 4.5’s efficiency is ridiculous > On Artificial Analysis, it costs just $0.31 per Intelligence Index task while delivering frontier intelligence > For comparison: > • Claude Fable 5 (max): $2.75 > • Claude Opus 4.8 (max): $1.80 > • GPT-5.6 Sol (max): $1.04 > • Kimi K3: $0.95 > • Grok 4.5: just $0.31 > That makes Grok 4.5 roughly: > • Nearly 9× cheaper than Claude Fable 5 > • Nearly 6× cheaper than Claude Opus 4.8 > • 3× cheaper per task than Kimi K3 > • 3.4× cheaper than GPT-5.6 Sol > On FrontierSWE, Grok 4.5 also used nearly 6× fewer total tokens per task than Fable 5 while still ranking near the very top > SpaceXAI has figured out how to deliver top-tier intelligence with extraordinary token efficiency > At this level of performance, no other frontier model is operating in the same cost-efficiency league > Grok 4.5 does not brute-force every problem with endless tokens and compute > It makes every token and every dollar work harder > https://x.com/XFreeze/status/2078171481415290947

相关动态

02

Kimi K3 在 SpreadsheetBench 2 上排名第一

Kimi K3 在 SpreadsheetBench 2 基准测试中排名第一,超越 Claude Fable 5,完成 34.8% 的 workflow 任务。该基准测试涉及 321 个专家策划任务,平均每个任务包含 11.8 个工作表和 593.5 个单元格更改,聚焦于完整的电子表格工作簿执行。

twitter关注列表2026-07-18#模型#技术突破#评测
观察
03

Mind Lab 开源长文本强化学习项目

Mind Lab 开源了一个长文本强化学习(RL)项目,支持 2M tokens 的长上下文。相比以往需要数千个 GPU 的 1M 上下文研究,该项目仅需 8 个 GPU 即可完成,大幅降低了超长上下文研究的算力门槛。

twitter关注列表2026-07-17#开源#技术突破#模型
观察
04

Kimi K3非对称竞争分析

Kimi K3在Arena前端/WebDev人类盲测中以1679 Elo排名第一,领先Fable 5(1631)和GPT-5.6 Sol(1618),但在综合智能指数中仅排第四(57.1分),落后Fable 5(59.9)和GPT-5.6 Sol(58.9)。K3定价15美元/百万输出tokens,远低于Fable 5的50美元,并计划7...

twitter关注列表2026-07-18#AI#模型#技术突破
观察