The Index / Weights & Biases / Nemotron 3.5 Lightning Nemotron 3.5 Lightning by Weights & Biases · family nemotron · listed Aug 2026R T W
Nemotron 3.5 Lightning is an MoE model built for fast, reliable agentic tasks across use cases such as financial services, cybersecurity, telecom, and retail.
Current rates USD per 1M tokens Input $0.10
Output $0.25
Cache read $0.05
Cache write —
Against the market Input price vs. maker flagships ($/1M tokens) Nemotron 3.5 Lightning Nemotron 3.5 Lightning: $0.10 per 1M tokens $0.10 DeepSeek V4 Pro DeepSeek V4 Pro: $0.435 per 1M tokens $0.435 Gemini 3.7 Flash Gemini 3.7 Flash: $0.75 per 1M tokens $0.75 Mistral Medium 3.5 Mistral Medium 3.5: $1.50 per 1M tokens $1.50 Grok 4.6 Grok 4.6: $2.00 per 1M tokens $2.00 Qwen3.8 Max Qwen3.8 Max: $2.00 per 1M tokens $2.00 Claude Opus 5 Claude Opus 5: $5.00 per 1M tokens $5.00 GPT-5.6 Sol GPT-5.6 Sol: $5.00 per 1M tokens $5.00 Output price vs. maker flagships ($/1M tokens) Nemotron 3.5 Lightning Nemotron 3.5 Lightning: $0.25 per 1M tokens $0.25 DeepSeek V4 Pro DeepSeek V4 Pro: $0.87 per 1M tokens $0.87 Gemini 3.7 Flash Gemini 3.7 Flash: $3.75 per 1M tokens $3.75 Grok 4.6 Grok 4.6: $6.00 per 1M tokens $6.00 Qwen3.8 Max Qwen3.8 Max: $6.00 per 1M tokens $6.00 Mistral Medium 3.5 Mistral Medium 3.5: $7.50 per 1M tokens $7.50 Claude Opus 5 Claude Opus 5: $25.00 per 1M tokens $25.00 GPT-5.6 Sol GPT-5.6 Sol: $30.00 per 1M tokens $30.00 On the record
Model id nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B
Context window 262K tokens
Max output 262K tokens
Input modalities text
Output modalities text
Reasoning yes
Tool calls yes
Open weights yes
Knowledge cutoff —
Released 2026-08-11
Last updated 2026-08-11 ← All Weights & Biases listings