Llama 3.2 1b Instruct
by Nvidia · listed Sep 2024TW
Open Llama instruction model for multilingual chat, reasoning, and coding
Current rates USD per 1M tokens
Inputfree
Outputfree
Cache read—
Cache write—
Against the market
On the record
- Model id
meta/llama-3.2-1b-instruct- Context window
- 128K tokens
- Max output
- 4K tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- no
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- Dec 2023
- Released
- 2024-09-18
- Last updated
- 2024-09-18
- Provider docs
- docs.api.nvidia.com/nim/
Back of the envelope
≈
Also listed at
- Inference$0.01 in · $0.01 out
- Cloudflare Workers AI$0.027 in · $0.201 out
- OpenRouter$0.027 in · $0.201 out
- Cloudflare AI Gateway$0.027 in · $0.201 out
- Kilo Gateway$0.027 in · $0.201 out
- Pioneer$0.10 in · $0.201 out