Model Details
Where is Nemotron-3 Ultra-550b-a55b a perfect fit?
- Complex mathematical, scientific, and technical reasoning
- Large-scale software engineering and codebase analysis
- Autonomous multi-step agents
- Tool calling and agentic workflow execution
- Long-context document and knowledge-base analysis
- Technical research and enterprise knowledge work
- Code generation, debugging, and refactoring
- Self-hosted deployments where open weights and infrastructure control are important
Quick Model Estimate
Your Nemotron-3 Ultra-550b-a55b Cost Estimate
๐ฐ Total Cost
โ
for 1000 input + 1000 output tokens
Cost Breakdown
Prices are watched for changes and checked against the provider's own page.
Pricing
|
Provider
โ
|
Modality
โ
|
Service Tier
โ
|
Input Price
(per 1M tokens) โ |
Output Price
(per 1M tokens) โ |
Cached Input
(per 1M tokens) โ |
Context Size
โ
|
View
|
|---|---|---|---|---|---|---|---|
Perplexity | Text | Standard | $0.2500 | $2.5000 | $0.2500 | 1,048,576 tokens | โ |
Other Models in the Nemotron-3 Family
FAQs about Nemotron-3 Ultra-550b-a55b
How much does Nemotron-3 Ultra-550b-a55b cost per 1M tokens?
Nemotron-3 Ultra-550b-a55b costs $0.25 per million input tokens, and $2.50 per million output tokens.
What does a typical workload cost with Nemotron-3 Ultra-550b-a55b?
1,000 requests of 2,000 input and 500 output tokens each โ 2,000,000 input and 500,000 output tokens in total โ costs $1.75 with Nemotron-3 Ultra-550b-a55b at its lowest rates. Use the calculator on this page for your own volumes.
Where can I use Nemotron-3 Ultra-550b-a55b?
Nemotron-3 Ultra-550b-a55b is available through Perplexity, from $0.25 per million input tokens.
What is Nemotron-3 Ultra-550b-a55b's context window?
Nemotron-3 Ultra-550b-a55b has a context window of 1,048,576 tokens.
