Llama 3.3 70b Instruct
by Nvidia · listed Nov 2024TW
Open Llama instruction model for multilingual chat, reasoning, and coding
Current rates USD per 1M tokens
Inputfree
Outputfree
Cache read—
Cache write—
Against the market
On the record
- Model id
meta/llama-3.3-70b-instruct- Context window
- 128K tokens
- Max output
- 4K tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- no
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- —
- Released
- 2024-11-26
- Last updated
- 2024-11-26
- Provider docs
- docs.api.nvidia.com/nim/
Back of the envelope
≈
Also listed at
- Llamafree in · free out
- NanoGPT$0.05 in · $0.23 out
- Meganova$0.10 in · $0.30 out
- OpenRouter$0.10 in · $0.32 out
- Eden AI$0.10 in · $0.32 out
- Kilo Gateway$0.10 in · $0.32 out
- Cortecs$0.129 in · $0.399 out
- LLM Gateway$0.13 in · $0.40 out
- Nebius Token Factory$0.13 in · $0.40 out
- IO.NET$0.13 in · $0.38 out
- Helicone$0.13 in · $0.39 out
- DevPass (LLM Gateway)$0.13 in · $0.40 out
- NovitaAI$0.135 in · $0.40 out
- Merge Gateway$0.22 in · $0.50 out
- Crusoe$0.25 in · $0.75 out