NVIDIA Nemotron 3 Nano 30B A3B vs gpt-oss-20b: Release Comparison
Comparison of the NVIDIA Nemotron 3 Nano 30B A3B and gpt-oss-20b releases: 2 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 4 models.
- For intelligence, gpt-oss-20b scores highest: gpt-oss-20b (low) at 10, against NVIDIA Nemotron 3 Nano 30B A3B (Reasoning) at 9 for NVIDIA Nemotron 3 Nano 30B A3B.
- For output speed, gpt-oss-20b is fastest: gpt-oss-20b (low) at 224 t/s, against NVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning) at 179 t/s for NVIDIA Nemotron 3 Nano 30B A3B.
- For cost per task, gpt-oss-20b is cheapest: gpt-oss-20b (high) at $0.01, against NVIDIA Nemotron 3 Nano 30B A3B (Reasoning) at $0.02 for NVIDIA Nemotron 3 Nano 30B A3B.
Releases compared
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost per Task (USD, Log Scale)
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indexes
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Further details
Input modality | Output modality | Weights | Provider Benchmarks | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| gpt-oss-20b (low) | 10 | - | - | - | - | - | - | - | $0.1 | $0.07 | $0.215 | - | - | - | - | - | 224 | 0.92s | 9.84s | 12.07s | - | 21B 3.6B active at inference time | 131k | Aug 2025 | May 2025 | Yes | Supports: text | Supports: text | +4 | ||||
| gpt-oss-20b (high) | 9 | 7 | 5 | 7 | 7 | 12 | 11 | $0.01 | $0.1 | $0.07 | $0.18 | - | $31 | 18k | 16k | 70M | 186 | 0.84s | 11.62s | 14.31s | 104.61s | 21B 3.6B active at inference time | 131k | Aug 2025 | May 2024 | Yes | Supports: text | Supports: text | +5 | ||||
| NVIDIA Nemotron 3 Nano 30B A3B (Reasoning) | 9 | 9 | 6 | 10 | - | 13 | 15 | $0.02 | $0.1 | $0.05 | $0.20 | - | $64 | 27k | 24k | 178M | 134 | 1.23s | 16.20s | 19.94s | 181.14s | 31.6B 3.6B active at inference time | 1M | Dec 2025 | - | Yes | Supports: text | Supports: text | |||||
| NVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning) | 7 | - | - | - | - | - | - | - | $0.1 | $0.05 | $0.20 | - | - | - | - | - | 179 | 0.95s | 0.95s | 3.74s | - | 31.6B 3.6B active at inference time | 1M | Dec 2025 | - | No | Supports: text | Supports: text | |||||