Monday, October 5, 2026 · No. 435Prices synced 20 hours ago
All·AI·Model

The AI Model Price Index

Llama 3.1 Nemotron Nano 8B v1

by Nvidia · family nemotron · listed Mar 2025RTW

Nemotron model for efficient reasoning, coding, and specialized AI agents

Current rates USD per 1M tokens

Inputfree
Outputfree
Cache read—
Cache write—

Against the market

Input price vs. maker flagships ($/1M tokens)
Llama 3.1 Nemotron Nano 8B v1Llama 3.1 Nemotron Nano 8B v1: free per 1M tokensfreeDeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.15 per 1M tokens$0.15Qwen3.8 Omni FlashQwen3.8 Omni Flash: $0.15 per 1M tokens$0.15Gemini 3.8 FlashGemini 3.8 Flash: $0.75 per 1M tokens$0.75GLM-5.3GLM-5.3: $1.40 per 1M tokens$1.40Claude Sonnet 5.5Claude Sonnet 5.5: $2.00 per 1M tokens$2.00GPT-6.1 SolGPT-6.1 Sol: $2.00 per 1M tokens$2.00Grok 4.7Grok 4.7: $2.00 per 1M tokens$2.00
Output price vs. maker flagships ($/1M tokens)
Llama 3.1 Nemotron Nano 8B v1Llama 3.1 Nemotron Nano 8B v1: free per 1M tokensfreeQwen3.8 Omni FlashQwen3.8 Omni Flash: $0.47 per 1M tokens$0.47DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.60 per 1M tokens$0.60Gemini 3.8 FlashGemini 3.8 Flash: $3.75 per 1M tokens$3.75GLM-5.3GLM-5.3: $4.40 per 1M tokens$4.40Grok 4.7Grok 4.7: $6.00 per 1M tokens$6.00Claude Sonnet 5.5Claude Sonnet 5.5: $10.00 per 1M tokens$10.00GPT-6.1 SolGPT-6.1 Sol: $10.00 per 1M tokens$10.00

On the record

Model id
nvidia/llama-3.1-nemotron-nano-8b-v1
Context window
131K tokens
Max output
16K tokens
Input modalities
text
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
yes
Knowledge cutoff
—
Released
2025-03-18
Last updated
2025-03-18

Back of the envelope

≈

← All Nvidia listings