Nemotron Cascade 2 30B A3B
Nemotron model for efficient reasoning, coding, and specialized AI agents.
- Capability
- —No ECI score yet
- Input
- —No listed price
- Output
- —No listed price
- Context
- 256K33K max output
What it costs
No per-token price is published for this model yet.
No per-token price is listed for this model on models.dev.
Independent benchmarks
Epoch AI has not published a Capabilities Index score for Nemotron Cascade 2 30B A3B yet.
The details
- Lab
- NVIDIA
- Released
- Mar 24, 2026
- Context window
- 256,000 tokens
- Max output
- 32,768 tokens
- Inputs
- Text
- Output
- Text
- Reasoning
- Always on
- Tool calling
- Yes
- Structured output
- No
- Weights
- Open weights
More from NVIDIA
About Nemotron Cascade 2 30B A3B
How much does Nemotron Cascade 2 30B A3B cost?
Nemotron Cascade 2 30B A3B has no published per-token API price on models.dev yet.
What is the context window of Nemotron Cascade 2 30B A3B?
Nemotron Cascade 2 30B A3B accepts up to 256,000 tokens per request and can write up to 32,768 tokens in one response.
How good is Nemotron Cascade 2 30B A3B?
Epoch AI has not published a Capabilities Index score for Nemotron Cascade 2 30B A3B yet.
Is Nemotron Cascade 2 30B A3B open source?
Yes. NVIDIA publishes the weights, so it can be downloaded and self-hosted.
What inputs does Nemotron Cascade 2 30B A3B support?
Nemotron Cascade 2 30B A3B accepts text and replies in text. It is a reasoning model, supports tool calling.
When was Nemotron Cascade 2 30B A3B released?
NVIDIA released Nemotron Cascade 2 30B A3B on Mar 24, 2026.
What are the best alternatives to Nemotron Cascade 2 30B A3B?
The closest current models from other labs on capability, price and release date are MiniMax-M2.7-highspeed (MiniMax, $0.60 / $2.40), Mercury Edit 2 (Inception, $0.25 / $0.75), GLM-5V-Turbo (Z.ai (Zhipu), $1.20 / $4.00) and Mistral Small 4 (Mistral AI, $0.15 / $0.60).