AI Research Atlas

Kimi k1.5

Moonshot AI · 20 January 2025

Multimodal long-CoT RL model matching o1 on math, code and vision benchmarks, with a paper detailing the recipe, released the same day as DeepSeek-R1.

Open-recipe account of scaling RL on LLMs, with 128K context scaling and online mirror descent policy optimization, and no MCTS, value functions or process reward models. It also uses 'long2short' distillation to short-CoT models. Weights not released.

Date
Monday, 20 January 2025
Lab
Moonshot AI
Kind
model
Access
app only

Figures

MeasureValueMeasured by
AIME 2024 (long-CoT)77.5
arXiv abstract
company
MATH-500 (long-CoT)96.2company
Codeforces (long-CoT)94th percentilecompany
AIME (short-CoT)60.8
vs GPT-4o and Claude 3.5 Sonnet baselines
company

Release date 2025-01-20 per Moonshot's GitHub/blog and Wikipedia; arXiv v1 is 2025-01-22. Benchmarks are company-reported. Overshadowed in coverage by DeepSeek-R1 on the same day.

Sources

  1. arxiv.org/abs/2501.12599
  2. github.com/MoonshotAI/Kimi-k1.5
  3. en.wikipedia.org/wiki/Kimi_(AI)

This record was checked against its sources on 6 October 2026. How we check

Related