Compare AIFind AIAI NewsAI How-To
About Us
PrivacyTermsFAQContactContact
AIB Inc.Company info
© 2026 AIB Inc.
OpenAI
OpenAI

GPT OSS 120B

Release Date 2025-08-05(Knowledge cutoff 2024-06-30)··131K IN · 131K OUTOpen Model(Apache 2.0)|API available
Overall1.5Top 92%Automation1.5Top 87%Office1.5Top 93%Coding1.5Top 94%

GPT-OSS-120B is OpenAI's first open-weight language model, featuring 117 billion total parameters in a Mixture-of-Experts architecture that activates just 5.1 billion per forward pass. Optimized to run on a single 80GB GPU with native MXFP4 quantization, it achieves near-parity with o4-mini on core reasoning benchmarks while supporting configurable reasoning depth, full chain-of-thought access, and native tool use including function calling and structured outputs. Released under the Apache 2.0 license, it brings frontier-level reasoning and agentic capabilities to a fully customizable, locally deployable model.

  1. 01ImagesGenerated images
  2. 02CostCost · Usage · Speed
  3. 03User CompareUser Comparisons
  4. 04Expert EvalExpert Performance Evaluation
  5. 05OpinionPublic Opinion
  6. 06Related Info

API cost

Market price (median)
$0.1IN$0.5OUT
Price range15 providers
$0.03~$0.35IN$0.17~$0.95OUT
Calculate Cost
IN
IN
$0.1-8.5%
$0.1-8.5%
OUT
OUT
$0.5-13.8%
$0.5-13.8%
9/2310/8
Calculate Cost
SourceOpenRouterLast updatedOct 8, 2026

Real-world usage

No.415 places0.3% share · 228B tokens/wk
472B
8/24
462B
8/31
549B
9/7
624B
9/14
631B
9/21
553B
9/28
228B
10/5
Top use cases
  • Summarization4.1%
  • Finance & trading2.4%
SourceOpenRouterLast updatedOct 8, 2026

Output Speed

Very fastper 2,000 tokens
First Output0.4s
total21.7s
94tok/s
First Output 0.4stotal: 21.7s
SourceOpenRouterLast updatedOct 8, 2026

User Comparisons

Writing Style
Answer length
Average
top 59%
Vocabulary diversity
Average
top 59%
Sentence rhythm
Average
top 41%
Language mixing
Very high
top 16%
List usage
Occasionally
15% of lines
Heading usage
Occasionally
7% of lines
Reasoning share
Lighter side
15% of total
Response Speed
instability ±241%
First response 7.5sTotal 26.2s
First response 7.5sTotal 26.2s
45 tok/s
19.5ms per token
SourceAIB Inc.Last updatedSep 19, 2026

Expert Performance Evaluation

Overall performance

MoreAll overall performance AI rankings
11.6Reasoning effort: High135280.8%

By effortLow 10.2High 11.6

About AA Intelligence IndexMore about AA Intelligence Index

Automation (agents)

MoreAll automation AI rankings
12.8%Reasoning effort: High−$2265.8%42 min

By effortLow 2.9%High 12.8%

About τ³-BankingMore about τ³-Banking

Documents and office work

MoreAll documents and office work AI rankings
52.0%Reasoning effort: High13524.4%58.2%44.4%21.5%28.3%

By effortLow 46.0%High 52.0%

About Long Context ReasoningMore about Long Context Reasoning

Research and fact-checking

MoreAll research and fact-checking AI rankings
138519.0%14.2%85.8%

About Arena Text FactualityMore about Arena Text Factuality

Coding

MoreAll coding AI rankings
0.0%Reasoning effort: High139057626.2%23.5%30.487.8%34.0%48.2%41.8%138518.7%1.41×11.0%

About Terminal-Bench 4.0More about Terminal-Bench 4.0

Writing and chat

MoreAll writing and chat AI rankings
1309132769.0%127813237.71 / 101286

About Arena Writing, Literature & LanguageMore about Arena Writing, Literature & Language

Math and science

MoreAll math and science AI rankings
19.6%Reasoning effort: High1.1%138178.2%88.9%93.493.4%25.0%1361138422.1%20.0%2.0%76.3%22.1

By effortLow 5.9%High 19.6%

About Humanity's Last ExamMore about Humanity's Last Exam

Specialized fields

1266132781.3%136213571367133992.5%85.6%86.0%90.0%90.0%64.4%73.4%55.5%92.0%85.5%95.6%51.8%79.1%73.0%64.0%80.0%30.0%11.0%74.0%8.9%74.0%59.0%47.0%94.0%55.0%73.5%1.0%45.2%4.7%74.0%84.2%69.8%46.4%13.1%87.7%76.8%56.6%75.7%

About Arena KoreanMore about Arena Korean

See each evaluation’s detail page for its source and last update·Data sources & removal requests

Related Info

OpenAI
PrivateAI Native
OpenAI
🇺🇸 San Francisco2015ARR $20B+$840B~3,500
GPT-6 Luna DecisionsGPT-6.1 SolGPT-6 SolGPT-6 LunaGPT Image 2.5 Flare