Should you use Nemotron 3 Ultra 550B A55B?
Nemotron 3 Ultra 550B A55B is listed for chat workloads and supports a 1.0M context window.
Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.
https://integrate.api.nvidia.com/v1 nvidia/nemotron-3-ultra-550b-a55b | Benchmark | Score | Metric | Date |
|---|---|---|---|
| SWE-Bench Verified | 70.7 | resolved | 2026-06-04 |
| SWE-Bench Multilingual | 67.7 | resolve rate | 2026-06-04 |
| Terminal-Bench | 56.4 | success rate | 2026-06-04 |
| GPQA | 87 | accuracy | 2026-06-04 |
Nemotron 3 Ultra 550B A55B is listed for chat workloads and supports a 1.0M context window.
Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.
We found 3 provider listings for nemotron-3-ultra-550b-a55b. Check model ID, quota, pricing, and API format before switching providers.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | Nemotron 3 Ultra 550B A55B | Free tier | 1.0M | Native | Varies |
| | NVIDIA: Nemotron 3 Ultra (free) | Free tier | 1.0M | OpenAI-style | 200 req/day (free tier) |
| | nvidia/nemotron-3-ultra-550b-a55b:free | Free tier | 1.0M | OpenAI-style | ~200 req/hr |
| Model | Provider | Context | Access |
|---|---|---|---|
| NVIDIA: Nemotron 3 Ultra (free) | OpenRouter | 1.0M | Free tier |
| nvidia/nemotron-3-ultra-550b-a55b:free | Kilo Code | 1.0M | Free tier |
| 01-ai/yi-large | NVIDIA NIM | 131K | Free tier |
| adept/fuyu-8b | NVIDIA NIM | 131K | Free tier |
| ai21labs/jamba-1.5-large-instruct | NVIDIA NIM | 131K | Free tier |
Nemotron 3 Ultra 550B A55B is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
Nemotron 3 Ultra 550B A55B is listed with free API access on NVIDIA NIM, subject to the provider's quota and account policy.
The model ID shown in this catalog is nvidia/nemotron-3-ultra-550b-a55b.
The listed context window is 1.0M tokens with up to 128K output tokens.
Free Nemotron 3 Ultra 550B A55B API.
For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.