ZMIME
NVIDIA · Released Mar 24, 2026

Nemotron Cascade 2 30B A3B

Nemotron model for efficient reasoning, coding, and specialized AI agents.

  • Open weights
  • Reasoning
  • Tool calling
Capability
—No ECI score yet
Input
—No listed price
Output
—No listed price
Context
256K33K max output
Pricing

What it costs

No per-token price is published for this model yet.

No per-token price is listed for this model on models.dev.

Capability

Independent benchmarks

Epoch AI has not published a Capabilities Index score for Nemotron Cascade 2 30B A3B yet.

Specifications

The details

Lab
NVIDIA
Released
Mar 24, 2026
Context window
256,000 tokens
Max output
32,768 tokens
Inputs
Text
Output
Text
Reasoning
Always on
Tool calling
Yes
Structured output
No
Weights
Open weights
Same lab

More from NVIDIA

Questions

About Nemotron Cascade 2 30B A3B

How much does Nemotron Cascade 2 30B A3B cost?

Nemotron Cascade 2 30B A3B has no published per-token API price on models.dev yet.

What is the context window of Nemotron Cascade 2 30B A3B?

Nemotron Cascade 2 30B A3B accepts up to 256,000 tokens per request and can write up to 32,768 tokens in one response.

How good is Nemotron Cascade 2 30B A3B?

Epoch AI has not published a Capabilities Index score for Nemotron Cascade 2 30B A3B yet.

Is Nemotron Cascade 2 30B A3B open source?

Yes. NVIDIA publishes the weights, so it can be downloaded and self-hosted.

What inputs does Nemotron Cascade 2 30B A3B support?

Nemotron Cascade 2 30B A3B accepts text and replies in text. It is a reasoning model, supports tool calling.

When was Nemotron Cascade 2 30B A3B released?

NVIDIA released Nemotron Cascade 2 30B A3B on Mar 24, 2026.

What are the best alternatives to Nemotron Cascade 2 30B A3B?

The closest current models from other labs on capability, price and release date are MiniMax-M2.7-highspeed (MiniMax, $0.60 / $2.40), Mercury Edit 2 (Inception, $0.25 / $0.75), GLM-5V-Turbo (Z.ai (Zhipu), $1.20 / $4.00) and Mistral Small 4 (Mistral AI, $0.15 / $0.60).