Qwen 3 8B
by NanoGPT · family qwen · listed Jan 2024TW
Qwen 3 8B is a 8B model. Supports switching between thinking and non thinking: trigger thinking with /think and /no_think anywhere in a prompt or system message to toggle chain-of-thought reasoning.
Current rates USD per 1M tokens
Input$0.47
Output$0.47
Cache read$0.235
Cache write—
Against the market
On the record
- Model id
qwen/qwen3-8b- Context window
- 41K tokens
- Max output
- 33K tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- no
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- —
- Released
- 2024-01-01
- Last updated
- 2024-01-01
- Provider docs
- docs.nano-gpt.com
Back of the envelope
≈
Also listed at
- SiliconFlow$0.06 in · $0.06 out
- SiliconFlow (China)$0.06 in · $0.06 out
- Alibaba (China)$0.072 in · $0.287 out
- Kilo Gateway$0.117 in · $0.455 out
- OpenRouter$0.117 in · $0.455 out
- Alibaba$0.18 in · $0.70 out
- Pioneer$0.20 in · $0.20 out