月之暗面发布 Kimi K1.5

长上下文推理新范式,国产推理模型的国际水平

月之暗面发布 Kimi K1.5,用强化学习训练出具备长上下文推理能力的模型,多项基准对齐或超越当时的国际头部推理模型,成为国产推理模型的代表作品。

时间2025 年 1 月 20 日 级别A · 行业级 组织月之暗面 / Moonshot AI 状态已核验 · 1 个来源
新月与展开的发光卷轴插画
Kimi K1.5 把长上下文与强化学习推理结合,定义了国产推理模型的新高度。 AI Chronicle

2025 年 1 月 20 日,月之暗面发布了 Kimi K1.5。对国产 AI 而言,这份发布有一个特别的分量:它证明中国的推理模型可以站到国际头部水平,而不只是「在国内领先」。K1.5 用强化学习训练推理能力,把「长上下文」这个老招牌和新焦点「推理」焊在了一起。

月之暗面的起点是长上下文。Kimi 靠「200 万字上下文」在国内脱颖而出,「读得长」是它的标签。但 2025 年的行业焦点已经转向推理——模型不仅要读得多,还要想得深。K1.5 要回答的问题是:一个以长上下文著称的团队,能不能在推理上也追平国际最强?

答案写在评测里。K1.5 在数学、编程、长文档推理等任务上对齐甚至超过当时的国际头部推理模型,尤其是超长上下文下的多步推理表现亮眼。它不是某一项的偏科生,而是把「读得长」和「想得深」结合起来——这正是 K1.5 独特的路线:长上下文 + RL 推理。

这条路线的意义不止于一家公司。它证明了国产模型可以在推理这个硬核赛道上成规模地追上国际头部,也为后来的 K2 系列铺好了路。当越来越多中国团队不再满足于「追国内」,而是对标世界最强,整个国产 AI 的水平线都在被抬高。

回看 K1.5,它值得被记住的是那个组合拳:在别人还在单点突破时,月之暗面把长上下文和推理这两块拼图拼到了一起,拼出了国际水平的成绩。它让「中国推理模型」从一句期许,变成了有作品支撑的结论。

On January 20, 2025 Moonshot released Kimi K1.5. For Chinese AI this release carried special weight: it proved a Chinese reasoning model could stand at the international-leader level, not merely "lead domestically". K1.5 trained reasoning with reinforcement learning, welding the old signature—long context—to the new focus of reasoning.

Moonshot's starting point was long context. Kimi stood out domestically on "2 million token context", and "reads long" was its label. But by 2025 the industry focus had shifted to reasoning: models must not only read much but think deep. K1.5 had to answer whether a team famous for long context could also match the world's strongest in reasoning.

The answer was in the evals. K1.5 matched or exceeded the then-international reasoning leaders on math, coding, and long-document reasoning, with particularly strong multi-step reasoning over very long contexts. It wasn't a one-subject specialist; it combined "reads long" with "thinks deep"—that was K1.5's distinctive route: long context plus RL reasoning.

The significance went beyond one company. It proved Chinese models could catch the international reasoning leaders at scale on this hard core, and paved the way for the K2 line. As more Chinese teams stop settling for "beating domestic rivals" and benchmark against the world's best, the whole level of Chinese AI rises.

Looking back, K1.5 deserves to be remembered for the combination: while others were making single-point breakthroughs, Moonshot fit two pieces—long context and reasoning—into one international-level result. It turned "Chinese reasoning model" from a hope into a conclusion backed by a real work.

展开完整事件档案人物、主题、模型与产品
人物
模型
产品
来源

原始资料

  1. 01Kimi K1.5 technical reportarXiv · paper

试试搜索