Thursday, September 17, 2026 · No. 417Prices synced 19 hours ago
All·AI·Model

The AI Model Price Index

Inkling

by Nvidia · family ling · listed Jul 2026RTVW

Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio

Current rates USD per 1M tokens

Inputfree
Outputfree
Cache read
Cache write

Against the market

Input price vs. maker flagships ($/1M tokens)
InklingInkling: free per 1M tokensfreeDeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.15 per 1M tokens$0.15Qwen3.8 FlashQwen3.8 Flash: $0.15 per 1M tokens$0.15Gemini 3.8 FlashGemini 3.8 Flash: $0.75 per 1M tokens$0.75GLM-5.2GLM-5.2: $1.40 per 1M tokens$1.40Grok 4.6Grok 4.6: $2.00 per 1M tokens$2.00Claude Fable 5.1Claude Fable 5.1: $10.00 per 1M tokens$10.00GPT-6 AstraGPT-6 Astra: $10.00 per 1M tokens$10.00
Output price vs. maker flagships ($/1M tokens)
InklingInkling: free per 1M tokensfreeQwen3.8 FlashQwen3.8 Flash: $0.47 per 1M tokens$0.47DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.60 per 1M tokens$0.60Gemini 3.8 FlashGemini 3.8 Flash: $3.75 per 1M tokens$3.75GLM-5.2GLM-5.2: $4.40 per 1M tokens$4.40Grok 4.6Grok 4.6: $6.00 per 1M tokens$6.00Claude Fable 5.1Claude Fable 5.1: $50.00 per 1M tokens$50.00GPT-6 AstraGPT-6 Astra: $50.00 per 1M tokens$50.00

On the record

Model id
thinkingmachines/inkling
Context window
1.0M tokens
Max output
16K tokens
Input modalities
text, image, audio
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
yes
Knowledge cutoff
Released
2026-07-15
Last updated
2026-07-15

Back of the envelope

Also listed at

← All Nvidia listings