Compare AIFind AIAI NewsAI How-To
About Us
PrivacyTermsFAQContactContact
AIB Inc.Company info
© 2026 AIB Inc.
Anthropic
Anthropic

Claude Sonnet 4

Release Date 2025-05-22(Knowledge cutoff 2025-01-31)··1M IN · 64K OUTProprietary Model|API available
Overall2.0Top 79%Automation2.5Top 52%Office2.0Top 83%Coding2.5Top 64%

Claude Sonnet 4 is Anthropic's balanced mid-tier model released alongside Opus 4 in May 2025, designed to combine strong coding and reasoning capabilities with computational efficiency. It achieves state-of-the-art 72.7% on SWE-bench while offering significantly lower cost and faster response times than Opus models. Key strengths include autonomous codebase navigation, reduced error rates in agent-driven workflows, and high reliability in following intricate instructions, making it a versatile choice for both routine and complex development tasks.

  1. 01ImagesGenerated images
  2. 02CostCost · Usage · Speed
  3. 03User CompareUser Comparisons
  4. 04Expert EvalExpert Performance Evaluation
  5. 05OpinionPublic Opinion
  6. 06Related Info

API cost

Official price
$3IN$15OUT
Calculate Cost
SourceAnthropic official docs

Output Speed

Averageper 2,000 tokens
First Output1.3s
total37.6s
55tok/s
First Output 1.3stotal: 37.6s
SourceOpenRouterLast updatedOct 8, 2026

User Comparisons

Response Speed
Total 6.9s
Total 6.9s
47 tok/s
SourceAIB Inc.Last updatedApr 3, 2026

Expert Performance Evaluation

Overall performance

MoreAll overall performance AI rankings
18.9Reasoning effort: Thinking140284.2%

By effortNone 16.6Thinking 18.9

About AA Intelligence IndexMore about AA Intelligence Index

Automation (agents)

MoreAll automation AI rankings
16.7%Reasoning effort: Thinking64.6%1h 15m33.1%43.9%

About τ³-BankingMore about τ³-Banking

Documents and office work

MoreAll documents and office work AI rankings
70.3%Reasoning effort: Thinking138161.2%35.0%46.9%44.5%

By effortNone 44.0%Thinking 70.3%

About Long Context ReasoningMore about Long Context Reasoning

Research and fact-checking

MoreAll research and fact-checking AI rankings
46.6%Reasoning effort: Budget 2K10.3%89.7%58.9

About Deep Research BenchMore about Deep Research Bench

Coding

MoreAll coding AI rankings
147465536.3%31.1%37.665.5%40.0%46.1%4.9%61.3%1445

About Arena CodingMore about Arena Coding

Writing and chat

MoreAll writing and chat AI rankings
1398142154.7%139614168.14 / 1072.4%1389

About Arena Writing, Literature & LanguageMore about Arena Writing, Literature & Language

Math and science

MoreAll math and science AI rankings
10.7%Reasoning effort: Thinking0.3%140477.7%71.1%74.374.3%1416141399.1%84.4%4.1%0.0%40.0%5.9%45.5%3.1%77.1%29.0

By effortNone 4.3%Thinking 10.7%

About Humanity's Last ExamMore about Humanity's Last Exam

Image understanding

MoreAll image understanding and analysis AI rankings
119134.0%37.0%44.8%

About Arena VisionMore about Arena Vision

Specialized fields

13381342143214331416140635.0%17.9%56.1%28.6%20.4%35.0%90.6%90.0%91.3%97.8%81.5%28.3%79.2%89.3%74.6%6.0%24.6%

About Arena KoreanMore about Arena Korean

See each evaluation’s detail page for its source and last update·Data sources & removal requests

Related Info

Anthropic
PrivateAI Native
Anthropic
🇺🇸 San Francisco2021ARR $30B+$380B~1,200
Claude Haiku 5.5Claude Sonnet 5.5Claude Opus 5.5Claude Fable 5.1Claude Opus 5