AI Research Atlas

MiMo-7B

Xiaomi (MiMo) · 30 April 2025

Xiaomi's first reasoning LLM: a 7B model pretrained for reasoning (25T tokens) and RL-tuned on 130K verifiable problems, claimed to beat OpenAI o1-mini.

Pretraining adds multi-token prediction and reasoning-dense data; RL uses a difficulty-based reward to handle sparse rewards. The base model is said to beat 32B models. MIT licence; paper arXiv 2505.07608 (2025-05-12).

Date
Wednesday, 30 April 2025
Lab
Xiaomi (MiMo)
Kind
open-weights
Access
open weights

Figures

MeasureValueMeasured by
Pretraining tokens25Tcompany
Verifiable RL problems130K
math and programming
company

Release 2025-04-30 per Wikipedia (HF repo 2025-04-29 UTC). Claims vs o1-mini are authors' own. Team led by Luo Fuli per secondary reporting (not verified here).

Sources

  1. arxiv.org/abs/2505.07608
  2. huggingface.co/XiaomiMiMo/MiMo-7B-RL
  3. en.wikipedia.org/wiki/Xiaomi_MiMo

This record was checked against its sources on 6 October 2026. How we check

Related