Signal Brief

Kimi K3 上线 AI/ML API 且性能超越 Claude Fable 5

AI/ML API 上线了 Kimi K3 模型,与 Claude Fable 5 均支持 1M-token 上下文窗口。在 vibe-coding 测试中,Kimi K3 以 1/4 的价格($3.11 对比 $12.23)在代码生成任务中表现优于 Claude Fable 5。

twitter关注列表 🚨 AI News | TestingCatalog (@testingcatalog) 发布 2026-07-16 收录 2026-07-17 观察

一句话判断

该模型在低成本下展示了极强的代码生成能力,值得关注其在游戏开发场景的实际表现。

核心信息

AI/ML API 上线了 Kimi K3 模型,与 Claude Fable 5 均支持 1M-token 上下文窗口。在 vibe-coding 测试中,Kimi K3 以 1/4 的价格($3.11 对比 $12.23)在代码生成任务中表现优于 Claude Fable 5。

原始内容

Kimi K3 is now live on day one on the AI/ML API, available alongside Claude Fable 5 under the same API key. > Moonshot lists game development as one of the workloads K3 is built for. > The AI/ML API team has run both models on the same vibe-coding test, with side-by-side outputs. > Both models carry a 1M-token context window. Who is the winner? ![photo](https://pbs.twimg.com/media/HNYxom_XEAAU-mY.jpg) > **引用原帖 AI/ML API (@aimlapi):** > Kimi K3 just outperformed Claude Fable 5 at a quarter of the price. > Kimi K3 $3.11 > Claude Fable 5 $12.23 > Same prompt: a one-shot MECCHA CHAMELEON, Steam hide-and-seek game. A white chameleon hides in a hand-drawn room, paints itself with the mouse to match the wall behind it, then survives three sweeps of a robot seeker. > We asked for a live pixel-diff match %, five procedural zones, synthesized sound, three scored rounds — one HTML file, no libraries. > Both models live on AI/ML API > https://x.com/aimlapi/status/2077898742179459274 🚨 AI News | TestingCatalog (@testingcatalog): AI/ML API allows users to use the same key, SDK, and billing. Models sit behind a single OpenAI-compatible endpoint, alongside 1000+ others, and switching between them requires a single string change. Test it out 👀 https://t.co/36GFFEah7w

相关动态

01

Kimi K3 在 SpreadsheetBench 2 上排名第一

Kimi K3 在 SpreadsheetBench 2 基准测试中排名第一,超越 Claude Fable 5,完成 34.8% 的 workflow 任务。该基准测试涉及 321 个专家策划任务,平均每个任务包含 11.8 个工作表和 593.5 个单元格更改,聚焦于完整的电子表格工作簿执行。

twitter关注列表2026-07-18#模型#技术突破#评测
观察
03

BestBlogs 早报 · 07-18

月之暗面发布Kimi K3,2.8万亿参数,896选16的Stable LatentMoE,上下文100万token,接近Fable-5但仍落后最强闭源模型,完整权重7月27日前开源;VentureBeat调查显示54%企业已发生AI代理安全事件;xAI开源Grok Build(84万行Rust代码)并残留上传用户代码痕迹;Cursor评...

twitter关注列表2026-07-17#AI#模型发布#开源
观察
04

Kimi K3 编码代理评测

Artificial Analysis发布Kimi K3在编码代理指数上的评测结果:得分57,排名第5,性能与GPT-5.6 Terra和GPT-5.5持平(57),超过Opus 4.8(55),略低于Grok 4.5(58)和Fable 5(59);成本平均3.18美元/任务,比GPT-5.6 Sol便宜55%。

twitter关注列表2026-07-17#AI模型#评测
观察