Friday, August 21, 2026 · No. 390Prices synced 4 hours ago
All·AI·Model

The AI Model Price Index

Gemini 3.5 Flash-Lite

by Venice AI · family gemini-flash-lite · listed Jul 2026RTV

Low-latency Gemini model for high-volume multimodal and agent workloads

Current rates USD per 1M tokens

Input$0.375
Output$3.13
Cache read$0.0375
Cache write

Against the market

Input price vs. maker flagships ($/1M tokens)
Gemini 3.5 Flash-LiteGemini 3.5 Flash-Lite: $0.375 per 1M tokens$0.375DeepSeek V4 ProDeepSeek V4 Pro: $0.435 per 1M tokens$0.435Gemini 3.7 FlashGemini 3.7 Flash: $0.75 per 1M tokens$0.75Mistral Medium 3.5Mistral Medium 3.5: $1.50 per 1M tokens$1.50Grok 4.6Grok 4.6: $2.00 per 1M tokens$2.00Qwen3.8 MaxQwen3.8 Max: $2.00 per 1M tokens$2.00Claude Opus 5Claude Opus 5: $5.00 per 1M tokens$5.00GPT-5.6 SolGPT-5.6 Sol: $5.00 per 1M tokens$5.00
Output price vs. maker flagships ($/1M tokens)
DeepSeek V4 ProDeepSeek V4 Pro: $0.87 per 1M tokens$0.87Gemini 3.5 Flash-LiteGemini 3.5 Flash-Lite: $3.13 per 1M tokens$3.13Gemini 3.7 FlashGemini 3.7 Flash: $3.75 per 1M tokens$3.75Grok 4.6Grok 4.6: $6.00 per 1M tokens$6.00Qwen3.8 MaxQwen3.8 Max: $6.00 per 1M tokens$6.00Mistral Medium 3.5Mistral Medium 3.5: $7.50 per 1M tokens$7.50Claude Opus 5Claude Opus 5: $25.00 per 1M tokens$25.00GPT-5.6 SolGPT-5.6 Sol: $30.00 per 1M tokens$30.00

On the record

Model id
gemini-3-5-flash-lite
Context window
1M tokens
Max output
66K tokens
Input modalities
text, image, audio, video
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
no
Knowledge cutoff
Mar 2026
Released
2026-07-09
Last updated
2026-07-21
Provider docs
docs.venice.ai

Back of the envelope

Also listed at

  • Neon$0.30 in · $2.50 out

← All Venice AI listings