Signal Brief

Zhipu AI 和 DeepSeek 探索定制 ASIC

Zhipu AI 和 DeepSeek 正在探索定制 ASIC,Zhipu GLM-5.2 近一周内使用量飙升 27 倍。定制芯片可降低功耗和每令牌成本,比通用 GPU 更适合大规模推理。然而,项目可能需 2 年多,且面临先进制造厂和高带宽内存的访问限制。

twitter关注列表 Rohan Paul (@rohanpaul_ai) 发布 2026-07-07 收录 2026-07-07 观察

一句话判断

中国 AI 公司向定制芯片的转移或将改变硬件软件一体化格局,值得关注。

核心信息

Zhipu AI 和 DeepSeek 正在探索定制 ASIC,Zhipu GLM-5.2 近一周内使用量飙升 27 倍。定制芯片可降低功耗和每令牌成本,比通用 GPU 更适合大规模推理。然而,项目可能需 2 年多,且面临先进制造厂和高带宽内存的访问限制。

原始内容

Per The Information, Zhipu AI is also (after DeepSeek) exploring a custom ASIC after GLM-5.2 usage reportedly jumped 27x in one week. A custom ASIC removes flexibility, but it can cut power draw and per-token cost. Nvidia GPUs are strong general-purpose machines, but inference at scale has different economics. A fixed model can run better on silicon designed around its own repeated operations. Zhipu has not chosen a partner, and the project may take more than 2 years. The pattern is now bigger than one Chinese lab or one model launch. Chinese AI companies are trying to make software, hardware, and deployment less separable. ![photo](https://pbs.twimg.com/media/HMosxAdakAADdwK.png) > **引用原帖 Rohan Paul (@rohanpaul_ai):** > DeepSeek is building an inference chip to cut dependence on Nvidia and Huawei in China’s $50B AI-chip market. > DeepSeek’s chip work is still early, with outside partners and private hiring of chip-design engineers. > The hard part is not drawing a chip, but making it at scale. Advanced foundries and high-bandwidth memory remain chokepoints because U.S. rules restrict Chinese access. > DeepSeek can still gain from a narrower chip built mainly for its own models. A custom inference chip could lower serving costs, reduce power needs, and tighten software-hardware control. > --- > reuters .com/world/china/chinas-deepseek-developing-its-own-ai-chip-sources-say-2026-07-07/ > https://x.com/rohanpaul_ai/status/2074519154292568401 Rohan Paul (@rohanpaul_ai): https://t.co/A02lJSHZy5

相关动态

02

Kimi K3 在 SpreadsheetBench 2 上排名第一

Kimi K3 在 SpreadsheetBench 2 基准测试中排名第一,超越 Claude Fable 5,完成 34.8% 的 workflow 任务。该基准测试涉及 321 个专家策划任务,平均每个任务包含 11.8 个工作表和 593.5 个单元格更改,聚焦于完整的电子表格工作簿执行。

twitter关注列表2026-07-18#模型#技术突破#评测
观察
03

Mind Lab 开源长文本强化学习项目

Mind Lab 开源了一个长文本强化学习(RL)项目,支持 2M tokens 的长上下文。相比以往需要数千个 GPU 的 1M 上下文研究,该项目仅需 8 个 GPU 即可完成,大幅降低了超长上下文研究的算力门槛。

twitter关注列表2026-07-17#开源#技术突破#模型
观察
04

Kimi K3非对称竞争分析

Kimi K3在Arena前端/WebDev人类盲测中以1679 Elo排名第一,领先Fable 5(1631)和GPT-5.6 Sol(1618),但在综合智能指数中仅排第四(57.1分),落后Fable 5(59.9)和GPT-5.6 Sol(58.9)。K3定价15美元/百万输出tokens,远低于Fable 5的50美元,并计划7...

twitter关注列表2026-07-18#AI#模型#技术突破
观察
05

二年级学生Jo Nagai发现蝴蝶可遗传记忆

日本东京都二年级学生Jo Nagai注意到养育的食蚜 caterpillars 在蝴蝶化后仍保持对薰衣草的回避行为,经Georgetown大学Entomologist Dr. Martha Weiss合作完成实验:70%训练过的蝴蝶及其后代均表现出对薰衣草的遗传性回避,记忆在全变态发育中保存并遗传。

twitter关注列表2026-07-18#研究#技术突破#信息
值得跟进