Llama 3.1 Nemotron Nano 8B v1
by Nvidia · family nemotron · listed Mar 2025RTW
Nemotron model for efficient reasoning, coding, and specialized AI agents
Current rates USD per 1M tokens
Inputfree
Outputfree
Cache read—
Cache write—
Against the market
Input price vs. maker flagships ($/1M tokens)Output price vs. maker flagships ($/1M tokens)On the record
- Model id
nvidia/llama-3.1-nemotron-nano-8b-v1- Context window
- 131K tokens
- Max output
- 16K tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- yes
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- —
- Released
- 2025-03-18
- Last updated
- 2025-03-18
← All Nvidia listings