Friday, August 21, 2026 · No. 390Prices synced 4 hours ago
All·AI·Model

The AI Model Price Index

nvidia-nemotron-3-super-120b-a12b

by Requesty · family nemotron · listed Mar 2026RT

NVIDIA Nemotron 3 Super is a hybrid Mixture-of-Experts (MoE) model engineered for highest compute efficiency and accuracy in multi-agent applications and specialized agentic systems. It is optimized to run many collaborating agents per application on a single GPU, delivering high accuracy for reasoning, tool use, and instruction following.

Current rates USD per 1M tokens

Input$0.09
Output$0.45
Cache read
Cache write

Against the market

Input price vs. maker flagships ($/1M tokens)
nvidia-nemotron-3-super-120b-a12bnvidia-nemotron-3-super-120b-a12b: $0.09 per 1M tokens$0.09DeepSeek V4 ProDeepSeek V4 Pro: $0.435 per 1M tokens$0.435Gemini 3.7 FlashGemini 3.7 Flash: $0.75 per 1M tokens$0.75Mistral Medium 3.5Mistral Medium 3.5: $1.50 per 1M tokens$1.50Grok 4.6Grok 4.6: $2.00 per 1M tokens$2.00Qwen3.8 MaxQwen3.8 Max: $2.00 per 1M tokens$2.00Claude Opus 5Claude Opus 5: $5.00 per 1M tokens$5.00GPT-5.6 SolGPT-5.6 Sol: $5.00 per 1M tokens$5.00
Output price vs. maker flagships ($/1M tokens)
nvidia-nemotron-3-super-120b-a12bnvidia-nemotron-3-super-120b-a12b: $0.45 per 1M tokens$0.45DeepSeek V4 ProDeepSeek V4 Pro: $0.87 per 1M tokens$0.87Gemini 3.7 FlashGemini 3.7 Flash: $3.75 per 1M tokens$3.75Grok 4.6Grok 4.6: $6.00 per 1M tokens$6.00Qwen3.8 MaxQwen3.8 Max: $6.00 per 1M tokens$6.00Mistral Medium 3.5Mistral Medium 3.5: $7.50 per 1M tokens$7.50Claude Opus 5Claude Opus 5: $25.00 per 1M tokens$25.00GPT-5.6 SolGPT-5.6 Sol: $30.00 per 1M tokens$30.00

On the record

Model id
nvidia-nemotron-3-super-120b-a12b
Context window
262K tokens
Max output
262K tokens
Input modalities
text
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
no
Knowledge cutoff
Released
2026-03-11
Last updated
2026-03-11

Back of the envelope

← All Requesty listings