Compare AIFind AIAI NewsAI How-To
About Us
PrivacyTermsFAQContactContact
AIB Inc.Company info
© 2026 AIB Inc.
xAI
xAI

Grok 4.20

Release Date 2026-03-09··2M IN · 2M OUTProprietary Model|API available
Overall2.5Top 49%
AutomationNo data
Office2.0Top 76%Coding3.0Top 39%

Grok 4.20 is xAI's newest flagship model released in February 2026, introducing a native 4-agent multi-agent architecture where specialized AI agents collaborate simultaneously on complex queries. It maintains a 2M-token context window — the largest among Western frontier models — and achieves a 65% reduction in hallucination rates through cross-agent verification. The model updates its capabilities weekly based on real-world usage and delivers fast direct answers at 232 tokens per second with 0.54-second time-to-first-token.

  1. 01ImagesGenerated images
  2. 02CostCost · Usage · Speed
  3. 03User CompareUser Comparisons
  4. 04Expert EvalExpert Performance Evaluation
  5. 05OpinionPublic Opinion
  6. 06Related Info

API cost

Provider price
$1.25IN$2.5OUT
Calculate Cost
SourceOpenRouterLast updatedSep 23, 2026

Output Speed

Fastper 2,000 tokens
First Output0.7s
total25.7s
80tok/s
First Output 0.7stotal: 25.7s
SourceOpenRouterLast updatedOct 8, 2026

User Comparisons

Response Speed
First response 1.0sTotal 13.9s
First response 1.0sTotal 13.9s
70 tok/s
SourceAIB Inc.Last updatedMay 15, 2026

Expert Performance Evaluation

Overall performance

MoreAll overall performance AI rankings
14.2Reasoning effort: None1475

About AA Intelligence IndexMore about AA Intelligence Index

Automation (agents)

MoreAll automation AI rankings
QCan it handle customer service by the rules?TAU2·Bottom 19%#65
59.9%Reasoning effort: None

About TAU2More about TAU2

Documents and office work

MoreAll documents and office work AI rankings
23.3%Reasoning effort: None1470

About Long Context ReasoningMore about Long Context Reasoning

Research and fact-checking

MoreAll research and fact-checking AI rankings
1452118961.7

About Arena Text FactualityMore about Arena Text Factuality

Coding

MoreAll coding AI rankings
150816.7%22.032.8%1504

About Arena CodingMore about Arena Coding

Writing and chat

MoreAll writing and chat AI rankings
1452148349.3%146314501449

About Arena Writing, Literature & LanguageMore about Arena Writing, Literature & Language

Math and science

MoreAll math and science AI rankings
27.9%Reasoning effort: None144977.6%14791455

About Humanity's Last ExamMore about Humanity's Last Exam

Specialized fields

14441478148714761498146810.0%

About Arena KoreanMore about Arena Korean

See each evaluation’s detail page for its source and last update·Data sources & removal requests

Related Info

xAI
PrivateAI Native
xAI
🇺🇸 Palo Alto2023ARR $500M+$250B~2,000
Grok 4.7Grok 4.6Grok 4.5Grok Build 0.1Grok Imagine Image Quality