OpenReasoning-Nemotron-32B is a reasoning model derived from Qwen2.5-32B-Instruct, post-trained for math, science, and code solution generation. Evaluated with up to 64K output tokens. Available in multiple sizes: 1.5B, 7B, 14B, and 32B.
Added Aug 21, 2025
Context Window
32.8K
Max Output
65.5K
Input Price (Auto)
$0.10/1M
Output Price (Auto)
$0.40/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare OpenReasoning Nemotron 32B with similar models from the same provider or model family.
Nemotron 3.5 Content Safety
nvidia/nemotron-3.5-content-safetyNemotron 3.5 Content Safety is a content moderation classifier that labels user messages and assistant responses as safe or unsafe. It supports a 131,072-token context window and optional reasoning.
Nvidia Nemotron 3.5 Lightning TEE
TEE/nemotron-3.5-lightningNVIDIA's open-weight 30B mixture-of-experts model with 3B active parameters for fast agentic workflows, coding, and tool use. Running inside a TEE (Trusted Execution Environment), with provider attestation support.
Nvidia Nemotron 3.5 Lightning
nvidia/nemotron-3.5-lightningNvidia Nemotron 3.5 Lightning is an open-weight 30B-A3B hybrid Mamba-Transformer mixture-of-experts model for high-throughput agentic, coding, tool-use, instruction-following, and long-context workloads. Thinking is disabled on this variant.
Nvidia Nemotron 3.5 Lightning Thinking
nvidia/nemotron-3.5-lightning:thinkingNvidia Nemotron 3.5 Lightning is an open-weight 30B-A3B hybrid Mamba-Transformer mixture-of-experts model for high-throughput agentic, coding, tool-use, instruction-following, and long-context workloads. This variant enables its reasoning trace.
Nvidia Nemotron 3 Ultra 550B
nvidia/nemotron-3-ultra-550b-a55bNvidia's Nemotron 3 Ultra 550B A55B model from the Nemotron 3 family. It uses a hybrid Mamba-Transformer MoE architecture. Provider-specific context limits vary, with the longest current route supporting up to 1M context.
Nvidia Nemotron 3 Ultra 550B Thinking
nvidia/nemotron-3-ultra-550b-a55b:thinkingNvidia's Nemotron 3 Ultra 550B A55B model from the Nemotron 3 family. It uses a hybrid Mamba-Transformer MoE architecture. Provider-specific context limits vary, with the longest current route supporting up to 1M context. Thinking enabled.