Thursday, September 17, 2026 · No. 417Prices synced 19 hours ago
All·AI·Model

The AI Model Price Index

Gemini 3.5 Flash Lite

by Vercel AI Gateway · family gemini-flash-lite · listed Jul 2026RTV

Low-latency Gemini model for high-volume multimodal and agent workloads

Current rates USD per 1M tokens

Input$0.30
Output$2.50
Cache read$0.03
Cache write

Against the market

Input price vs. maker flagships ($/1M tokens)
DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.15 per 1M tokens$0.15Qwen3.8 FlashQwen3.8 Flash: $0.15 per 1M tokens$0.15Gemini 3.5 Flash LiteGemini 3.5 Flash Lite: $0.30 per 1M tokens$0.30Gemini 3.8 FlashGemini 3.8 Flash: $0.75 per 1M tokens$0.75GLM-5.2GLM-5.2: $1.40 per 1M tokens$1.40Grok 4.6Grok 4.6: $2.00 per 1M tokens$2.00Claude Fable 5.1Claude Fable 5.1: $10.00 per 1M tokens$10.00GPT-6 AstraGPT-6 Astra: $10.00 per 1M tokens$10.00
Output price vs. maker flagships ($/1M tokens)
Qwen3.8 FlashQwen3.8 Flash: $0.47 per 1M tokens$0.47DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.60 per 1M tokens$0.60Gemini 3.5 Flash LiteGemini 3.5 Flash Lite: $2.50 per 1M tokens$2.50Gemini 3.8 FlashGemini 3.8 Flash: $3.75 per 1M tokens$3.75GLM-5.2GLM-5.2: $4.40 per 1M tokens$4.40Grok 4.6Grok 4.6: $6.00 per 1M tokens$6.00Claude Fable 5.1Claude Fable 5.1: $50.00 per 1M tokens$50.00GPT-6 AstraGPT-6 Astra: $50.00 per 1M tokens$50.00

On the record

Model id
google/gemini-3.5-flash-lite
Context window
1M tokens
Max output
65K tokens
Input modalities
text, image, pdf
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
no
Knowledge cutoff
Mar 2026
Released
2026-07-21
Last updated
2026-07-21

Back of the envelope

Also listed at

← All Vercel AI Gateway listings