Tuesday, August 25, 2026 · No. 394Prices synced 5 hours ago
All·AI·Model

The AI Model Price Index

Qwen3.5 4B

by EmpirioLabs AI · listed Mar 2026RTVW

Qwen3.5 4B is a low-cost multimodal reasoning model with 256K context, image and video input, function tools, and structured output.

Current rates USD per 1M tokens

Input$0.04
Output$0.07
Cache read$0.02
Cache write

Against the market

Input price vs. maker flagships ($/1M tokens)
Qwen3.5 4BQwen3.5 4B: $0.04 per 1M tokens$0.04DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.14 per 1M tokens$0.14Gemini Flash LatestGemini Flash Latest: $0.75 per 1M tokens$0.75Mistral Medium 3.5Mistral Medium 3.5: $1.50 per 1M tokens$1.50Grok 4.6Grok 4.6: $2.00 per 1M tokens$2.00Qwen3.8 MaxQwen3.8 Max: $2.00 per 1M tokens$2.00Claude Opus 5Claude Opus 5: $5.00 per 1M tokens$5.00GPT-5.6 SolGPT-5.6 Sol: $5.00 per 1M tokens$5.00
Output price vs. maker flagships ($/1M tokens)
Qwen3.5 4BQwen3.5 4B: $0.07 per 1M tokens$0.07DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.28 per 1M tokens$0.28Gemini Flash LatestGemini Flash Latest: $3.75 per 1M tokens$3.75Grok 4.6Grok 4.6: $6.00 per 1M tokens$6.00Qwen3.8 MaxQwen3.8 Max: $6.00 per 1M tokens$6.00Mistral Medium 3.5Mistral Medium 3.5: $7.50 per 1M tokens$7.50Claude Opus 5Claude Opus 5: $25.00 per 1M tokens$25.00GPT-5.6 SolGPT-5.6 Sol: $30.00 per 1M tokens$30.00

On the record

Model id
qwen3-5-4b
Context window
262K tokens
Max output
33K tokens
Input modalities
text, image, video
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
yes
Knowledge cutoff
Released
2026-03-02
Last updated
2026-03-02
Provider docs
docs.empiriolabs.ai

Back of the envelope

← All EmpirioLabs AI listings