NVIDIA logo

NVIDIA

Nemotron 3 Super

AdvancedOpen-weightAvailable in FlowHunt

Nemotron 3 Super is an advanced AI model suited for demanding production workloads, developed by NVIDIA and released in March 2026. As an open-weight model, its trained weights are publicly available — you can self-host it, fine-tune it on proprietary data, or run it on-premise without API dependencies. The 1M-token context window is large enough to hold entire codebases, lengthy technical documents, or extended agentic sessions without truncation.

On the FlowHunt AI Leaderboard it is one of 24 tracked models, evaluated across 4 benchmarks. Standout scores: 18.3% on HLE (#6 of 6); 36.0 pts on AA Intelligence Index (#8 of 9); 79.2% on GPQA Diamond (#10 of 14).

Nemotron 3 Super was released on March 11, 2026 as NVIDIA’s mid-tier open-weight model, using 12B active parameters from a 120B total MoE pool. Super was designed for organisations running existing NVIDIA H100 and H200 GPU clusters who wanted strong inference performance without the hardware upgrade to Blackwell. With MMLU-Pro at 80.1% and SWE-bench Verified at 65.8%, it provided frontier-adjacent performance on the GPU generation that most enterprise customers had already deployed. Super fits the recurring pattern of NVIDIA releasing paired model sizes — one for next-generation hardware buyers, one for current installed base.

Note: 2.2x throughput vs GPT-OSS-120B. Reproducible eval suite. 1M context.

Released
March 2026
Parameters
12B/120B MoE
Context
1M tokens
Weights
Open
Tier
Advanced
Provider
NVIDIA

Nemotron 3 Super Benchmark Rankings

Rank among all 24 models tracked on this leaderboard. SR = self-reported · 3P / AA / C = independently verified.

BenchmarkScoreRankSourceWhat it measures
SWE-bench Verified60.5%11 of 13SRReal GitHub issue resolution (500 verified issues)
GPQA Diamond79.2%10 of 14SRGraduate-level Google-proof science Q&A
HLE18.3%6 of 6SRHumanity's Last Exam (50+ STEM disciplines)
AA Intelligence Index36.0 pts8 of 9AAArtificial Analysis composite score (independent)

Detailed Benchmark Scores

SWE-V SR
60.5%
Real GitHub issue resolution (500 verified issues)
GPQA ◇ SR
79.2%
Graduate-level Google-proof science Q&A
HLE SR
18.3%
Humanity's Last Exam (50+ STEM disciplines)
AA Idx AA
36.0 pts
Artificial Analysis composite score (independent)

Nemotron 3 Super vs. Advanced Peers

Head-to-head benchmark comparison with other Advanced-tier models. Higher is better for all metrics.

BenchmarkNemotron 3 SuperClaude Sonnet 4.6GPT-4.1
SWE-V60.5%79.6%54.6%
GPQA ◇79.2%
HLE18.3%
AA Idx36.0 pts

About NVIDIA

NVIDIA Founded 1993 · Santa Clara, CA
Website →

NVIDIA Corporation designs and sells graphics processing units (GPUs) and system-on-chip units that have become the dominant hardware platform for AI model training and inference globally, with an estimated 80-to-95-percent market share in datacentre AI accelerators. Founded in 1993 by Jensen Huang, Chris Malachowsky, and Curtis Priem, NVIDIA has expanded from its original gaming-graphics focus into a full-stack AI computing company. Its Nemotron model family reflects a vertical integration strategy: by releasing open-weight models optimised specifically for NVIDIA hardware, the company aims to demonstrate the value of its Blackwell-generation GPU clusters while shaping the open-weight ecosystem around NVIDIA-optimised software stacks.

Use Cases

What to Use Nemotron 3 Super For

Software Engineering
Scientific & Technical Reasoning
Self-Hosted & Private Deployments
Long-Document & Codebase Analysis

Strengths & Limitations

Strengths

  • Solid coding assistance — 60% SWE-bench Verified
  • Advanced scientific Q&A — 79% GPQA Diamond
  • Self-hosted & data-private deployments (open-weight)
  • Very long documents and large codebases (1M context)

Browse the Full AI Model Leaderboard

Frequently asked questions

Learn more

Nemotron 3 Ultra — Benchmarks, Capabilities, and Use Cases
Nemotron 3 Ultra — Benchmarks, Capabilities, and Use Cases

Nemotron 3 Ultra — Benchmarks, Capabilities, and Use Cases

Nemotron 3 Ultra is a open-weight AI model from NVIDIA (55B/550B MoE), classified as Frontier tier with a 262K tokens context window. It is available directly i...

3 min read
MiniMax M3 — Benchmarks, Capabilities, and Use Cases
MiniMax M3 — Benchmarks, Capabilities, and Use Cases

MiniMax M3 — Benchmarks, Capabilities, and Use Cases

MiniMax M3 is a open-weight AI model from MiniMax (MoE (undisclosed)), classified as Frontier tier with a 1M tokens context window. It is available directly in ...

4 min read
Mistral Large 3 — Benchmarks, Capabilities, and Use Cases
Mistral Large 3 — Benchmarks, Capabilities, and Use Cases

Mistral Large 3 — Benchmarks, Capabilities, and Use Cases

Mistral Large 3 is a open-weight AI model from Mistral (41B/675B MoE), classified as Advanced tier with a 256K tokens context window. It is available directly i...

4 min read