Model Details
Owner
NVIDIA
Released
Jun 4, 2026
Status
Active
Context Length
1,048,576
Max Output
NA
Provider
Perplexity
Where is nemotron-3-ultra-550b-a55b a perfect fit?
NVIDIA Nemotron 3 Ultra 550B A55B is NVIDIA's largest Nemotron 3 model, built for advanced reasoning, agentic workflows, software engineering, and long-context analysis. It uses a hybrid Mamba-Transformer Mixture-of-Experts architecture with 550B total parameters and 55B active parameters per token.
- Complex mathematical, scientific, and technical reasoning
- Large-scale software engineering and codebase analysis
- Autonomous multi-step agents
- Tool calling and agentic workflow execution
- Long-context document and knowledge-base analysis
- Technical research and enterprise knowledge work
- Code generation, debugging, and refactoring
- Self-hosted deployments where open weights and infrastructure control are important
- Complex mathematical, scientific, and technical reasoning
- Large-scale software engineering and codebase analysis
- Autonomous multi-step agents
- Tool calling and agentic workflow execution
- Long-context document and knowledge-base analysis
- Technical research and enterprise knowledge work
- Code generation, debugging, and refactoring
- Self-hosted deployments where open weights and infrastructure control are important
Quick Model Estimate
Your GPT-5 Cost Estimate
💰 Total Cost
USD 3.00
for 1000 input + 1000 output tokens
📥 Input (1000 × $0.250000)
USD 1.5000
📤 Output (1000 × $2.500000)
USD 1.5000
Cost Breakdown
📥 Input 50%
📤 Output 50%
Prices updated daily from official provider data.
Pricing
|
Provider
↕
|
Modality
↕
|
Tier
↕
|
Input Price
(per 1M tokens) ↕ |
Output Price
(per 1M tokens) ↕ |
Context Size
↕
|
View
|
|---|---|---|---|---|---|---|
Perplexity
|
Text | Standard | $0.2500 | $2.5000 | 1,048,576 tokens | → |
