Llama 3.1 8B Instruct
by Neon · family llama · listed Jul 2024TW
Meta's compact open-weight Llama 3.1 model for fast, low-cost text generation
Current rates USD per 1M tokens
Input$0.15
Output$0.45
Cache read—
Cache write—
Against the market
Input price vs. maker flagships ($/1M tokens)Output price vs. maker flagships ($/1M tokens)On the record
- Model id
meta-llama-3-1-8b-instruct- Context window
- 131K tokens
- Max output
- 8K tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- no
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- Dec 2023
- Released
- 2024-07-23
- Last updated
- 2024-07-23
← All Neon listings