Thursday, October 1, 2026 · No. 431Prices synced 12 hours ago
All·AI·Model

The AI Model Price Index

Qwen3.8 2.4T A95B (NVFP4)

by RunInfra · family qwen · listed Aug 2026RTW

Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows

Current rates USD per 1M tokens

Input$2.00
Output$6.00
Cache read$0.20
Cache write—

Against the market

Input price vs. maker flagships ($/1M tokens)
DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.15 per 1M tokens$0.15Qwen3.8 Omni FlashQwen3.8 Omni Flash: $0.15 per 1M tokens$0.15Gemini 3.8 FlashGemini 3.8 Flash: $0.75 per 1M tokens$0.75GLM-5.3GLM-5.3: $1.40 per 1M tokens$1.40Qwen3.8 2.4T A95B (NVFP4)Qwen3.8 2.4T A95B (NVFP4): $2.00 per 1M tokens$2.00Claude Sonnet 5.5Claude Sonnet 5.5: $2.00 per 1M tokens$2.00GPT-6.1 SolGPT-6.1 Sol: $2.00 per 1M tokens$2.00Grok 4.7Grok 4.7: $2.00 per 1M tokens$2.00
Output price vs. maker flagships ($/1M tokens)
Qwen3.8 Omni FlashQwen3.8 Omni Flash: $0.47 per 1M tokens$0.47DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.60 per 1M tokens$0.60Gemini 3.8 FlashGemini 3.8 Flash: $3.75 per 1M tokens$3.75GLM-5.3GLM-5.3: $4.40 per 1M tokens$4.40Qwen3.8 2.4T A95B (NVFP4)Qwen3.8 2.4T A95B (NVFP4): $6.00 per 1M tokens$6.00Grok 4.7Grok 4.7: $6.00 per 1M tokens$6.00Claude Sonnet 5.5Claude Sonnet 5.5: $10.00 per 1M tokens$10.00GPT-6.1 SolGPT-6.1 Sol: $10.00 per 1M tokens$10.00

On the record

Model id
Inferact/Qwen3.8-2.4T-A95B-NVFP4
Context window
262K tokens
Max output
33K tokens
Input modalities
text
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
yes
Knowledge cutoff
—
Released
2026-08-12
Last updated
2026-08-12
Provider docs
runinfra.ai/docs

Back of the envelope

≈

← All RunInfra listings