Compare AIFind AIAI NewsAI How-To
About Us
PrivacyTermsFAQContactContact
AIB Inc.Company info
© 2026 AIB Inc.
NVIDIA
NVIDIA

Nemotron 3 Super

Release Date 2026-03-11(Knowledge cutoff 2026-02-01)··262K IN · 1M OUTOpen Model|API available
Overall1.5Top 89%Automation2.0Top 75%Office1.5Top 89%Coding1.5Top 93%

Nemotron 3 Super is NVIDIA's open hybrid Mamba-Transformer MoE model with 120 billion total parameters, activating just 12 billion for maximum compute efficiency. Its hybrid architecture integrates Mamba layers for sequence efficiency with Transformer layers for precision reasoning, delivering over 5× throughput compared to its predecessor. With a native 1M-token context window and NVFP4 precision optimized for Blackwell GPUs, it scores 85.6% on PinchBench — the best among open models — making it well-suited for complex multi-agent applications, software development, and agentic reasoning.

  1. 01ImagesGenerated images
  2. 02CostCost · Usage · Speed
  3. 03User CompareUser Comparisons
  4. 04Expert EvalExpert Performance Evaluation
  5. 05OpinionPublic Opinion
  6. 06Related Info

API cost

Market price (median)
$0.0825IN$0.425OUT
Price range2 providers
$0.08~$0.085IN$0.4~$0.45OUT
Calculate Cost
IN
IN
$0.0825-2.9%
$0.0825-2.9%
OUT
OUT
$0.425+6.2%
$0.425+6.2%
9/2310/8
Calculate Cost
SourceOpenRouterLast updatedSep 25, 2026

Real-world usage

No.3511 places0.4% share · 264B tokens/wk
460B
8/24
321B
8/31
352B
9/7
330B
9/14
380B
9/28
264B
10/5
SourceOpenRouterLast updatedOct 8, 2026

Output Speed

Very slowper 2,000 tokens
First Output1.1s
total106s
19tok/s
First Output 1.1stotal: 106s
SourceOpenRouterLast updatedOct 8, 2026

User Comparisons

Writing Style
Answer length
Shorter side
top 63%
Vocabulary diversity
Plainer side
top 74%
Sentence rhythm
Varied side
top 24%
Language mixing
High
top 21%
List usage
Often
31% of lines
Heading usage
Occasionally
10% of lines
Reasoning share
Lighter side
10% of total
Lies (Hallucination)
Honest answers
50%
Correct 0Honest don't-know 2Lies 2
Lies without doubt
2 questions
Response Speed
instability ±294%
First response 12.0sTotal 28.4s
First response 12.0sTotal 28.4s
33 tok/s
18.4ms per token
User Reviews
  • •

    The output speed is super fast and the content is accurate too. What's especially unique is that it's learning about itself.4 months ago

SourceAIB Inc.Last updatedSep 9, 2026

Expert Performance Evaluation

Overall performance

MoreAll overall performance AI rankings
12.8Reasoning effort: Thinking1361

About AA Intelligence IndexMore about AA Intelligence Index

Automation (agents)

MoreAll automation AI rankings
10.3%Reasoning effort: Thinking67.8%

About τ³-BankingMore about τ³-Banking

Documents and office work

MoreAll documents and office work AI rankings
65.7%Reasoning effort: Thinking1350

About Long Context ReasoningMore about Long Context Reasoning

Research and fact-checking

MoreAll research and fact-checking AI rankings
139031.3%22.8%

About Arena Text FactualityMore about Arena Text Factuality

Coding

MoreAll coding AI rankings
0.0%Reasoning effort: Thinking140721438.6%28.8%37.736.2%38.0%1404

About Terminal-Bench 4.0More about Terminal-Bench 4.0

Writing and chat

MoreAll writing and chat AI rankings
1325135171.5%130313421316

About Arena Writing, Literature & LanguageMore about Arena Writing, Literature & Language

Math and science

MoreAll math and science AI rankings
20.8%Reasoning effort: Thinking3.1%137380.0%13681399

About Humanity's Last ExamMore about Humanity's Last Exam

Specialized fields

137813931351136687.0%64.1%63.2%65.1%93.0%83.0%90.9%59.3%41.7%85.0%66.0%96.7%20.5%17.3%39.0%8.5%77.0%66.0%57.0%80.0%53.0%82.3%50.3%60.0%75.0%50.4%54.4%82.5%61.1%73.6%

About Arena Hard PromptsMore about Arena Hard Prompts

See each evaluation’s detail page for its source and last update·Data sources & removal requests

Related Info

NVIDIA
Public
NVIDIA
🇺🇸 Santa Clara1993$5.8T~36,000 (total)
Nemotron 3 UltraNemotron 3 Nano OmniNemotron 3 Nano 30B A3B