Multi-agent parallel reasoning from xAI's flagship tier
Grok 4 Heavy is the highest tier of xAI's Grok 4 family. Unlike the standard Grok 4, Heavy runs multiple internal agents in parallel — each working the problem from a different angle — then compares, synthesizes, and votes internally before answering. xAI positions it for mathematics, hard reasoning, and scientific analysis rather than quick Q&A.
Who it's for: Researchers, mathematicians, and analysts whose workday includes genuinely hard reasoning where a wrong answer costs more than a slow one.
4-8 internal agents tackle the problem in parallel with divergent strategies, then debate and selectively share findings.
Won or tied first place on Humanity's Last Exam (50.7%) and AIME 2025 math at launch among public reasoning models.
Responses can spend several minutes of extra test-time compute on heavy reasoning when the query warrants it.
Ships inside SuperGrok subscriptions on grok.com and the Grok iOS/Android apps, with DeepSearch built in.
The brute-force compute champion — best when the answer has to be right, not fast or cheap.