AI Research Atlas

Grok-1.5

xAI · 28 March 2024

Grok-1.5 raises context from 8K to 128K tokens and lifts math and code scores sharply over Grok-1.

128K context with perfect needle-in-a-haystack retrieval across the window (xAI test), MATH 23.9% to 50.6% and HumanEval 63.2% to 74.1%. Still behind the best closed models on knowledge benchmarks.

Date
Thursday, 28 March 2024
Lab
xAI
Kind
model
Access
app only

Figures

MeasureValueMeasured by
MATH50.6%
vs 23.9% for Grok-1
company
HumanEval74.1%
vs 63.2% for Grok-1
company
MMLU81.3%company
GSM8K90%company

Announced for early testers on X first; broader rollout to X Premium subscribers followed (exact rollout date not confirmed here; Wikipedia gives 2024-05-15).

Sources

  1. x.ai/news/grok-1.5

This record was checked against its sources on 6 October 2026. How we check

Related