Thursday, September 3, 2026 · No. 403Prices synced 3 hours ago
All·AI·Model

The AI Model Price Index

Gemini 3.8 Flash (Google Vertex AI)

by LLM Gateway · family gemini · listed Sep 2026RTV

Fast Gemini model balancing multimodal reasoning, tool use, and cost

Current rates USD per 1M tokens

Input$0.75
Output$3.75
Cache read$0.075
Cache write$0.0833

Against the market

Input price vs. maker flagships ($/1M tokens)
DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.14 per 1M tokens$0.14Qwen3.8 FlashQwen3.8 Flash: $0.15 per 1M tokens$0.15Gemini 3.8 Flash (Google Vertex AI)Gemini 3.8 Flash (Google Vertex AI): $0.75 per 1M tokens$0.75Gemini 3.8 FlashGemini 3.8 Flash: $0.75 per 1M tokens$0.75GLM-5.2GLM-5.2: $1.40 per 1M tokens$1.40Grok 4.6Grok 4.6: $2.00 per 1M tokens$2.00GPT-5.6 SolGPT-5.6 Sol: $4.00 per 1M tokens$4.00Claude Fable 5.1Claude Fable 5.1: $10.00 per 1M tokens$10.00
Output price vs. maker flagships ($/1M tokens)
DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.28 per 1M tokens$0.28Qwen3.8 FlashQwen3.8 Flash: $0.47 per 1M tokens$0.47Gemini 3.8 Flash (Google Vertex AI)Gemini 3.8 Flash (Google Vertex AI): $3.75 per 1M tokens$3.75Gemini 3.8 FlashGemini 3.8 Flash: $3.75 per 1M tokens$3.75GLM-5.2GLM-5.2: $4.40 per 1M tokens$4.40Grok 4.6Grok 4.6: $6.00 per 1M tokens$6.00GPT-5.6 SolGPT-5.6 Sol: $20.00 per 1M tokens$20.00Claude Fable 5.1Claude Fable 5.1: $50.00 per 1M tokens$50.00

On the record

Model id
google-vertex/gemini-3.8-flash
Context window
1.0M tokens
Max output
66K tokens
Input modalities
text, image, audio
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
no
Knowledge cutoff
Released
2026-09-02
Last updated
2026-09-02
Provider docs
llmgateway.io/docs

Back of the envelope

Also listed at

← All LLM Gateway listings