Monday, October 5, 2026 · No. 435Prices synced 7 hours ago
All·AI·Model

The AI Model Price Index

Gemini 3.5 Flash-Lite

by Venice AI · family gemini-flash-lite · listed Jul 2026RTV

Low-latency Gemini model for high-volume multimodal and agent workloads

Current rates USD per 1M tokens

Input$0.375
Output$3.13
Cache read$0.0375
Cache write—

Against the market

Input price vs. maker flagships ($/1M tokens)
DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.15 per 1M tokens$0.15Qwen3.8 Omni FlashQwen3.8 Omni Flash: $0.15 per 1M tokens$0.15Gemini 3.5 Flash-LiteGemini 3.5 Flash-Lite: $0.375 per 1M tokens$0.375Gemini 3.8 FlashGemini 3.8 Flash: $0.75 per 1M tokens$0.75GLM-5.3GLM-5.3: $1.40 per 1M tokens$1.40Claude Sonnet 5.5Claude Sonnet 5.5: $2.00 per 1M tokens$2.00GPT-6.1 SolGPT-6.1 Sol: $2.00 per 1M tokens$2.00Grok 4.7Grok 4.7: $2.00 per 1M tokens$2.00
Output price vs. maker flagships ($/1M tokens)
Qwen3.8 Omni FlashQwen3.8 Omni Flash: $0.47 per 1M tokens$0.47DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.60 per 1M tokens$0.60Gemini 3.5 Flash-LiteGemini 3.5 Flash-Lite: $3.13 per 1M tokens$3.13Gemini 3.8 FlashGemini 3.8 Flash: $3.75 per 1M tokens$3.75GLM-5.3GLM-5.3: $4.40 per 1M tokens$4.40Grok 4.7Grok 4.7: $6.00 per 1M tokens$6.00Claude Sonnet 5.5Claude Sonnet 5.5: $10.00 per 1M tokens$10.00GPT-6.1 SolGPT-6.1 Sol: $10.00 per 1M tokens$10.00

On the record

Model id
gemini-3-5-flash-lite
Context window
1M tokens
Max output
66K tokens
Input modalities
text, image, audio, video
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
no
Knowledge cutoff
Mar 2026
Released
2026-07-09
Last updated
2026-07-21
Provider docs
docs.venice.ai

Back of the envelope

≈

Also listed at

  • Neon$0.30 in · $2.50 out

← All Venice AI listings