Should you use Nemotron Mini 4B Instruct?
Nemotron Mini 4B Instruct is listed for chat workloads and supports a 128K context window.
Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.
https://integrate.api.nvidia.com/v1 nvidia/nemotron-mini-4b-instruct Nemotron Mini 4B Instruct is listed for chat workloads and supports a 128K context window.
Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.
We only found this Nemotron Mini 4B Instruct listing on NVIDIA NIM in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | Nemotron Mini 4B Instruct | Free tier | 128K | Native | Varies |
| Model | Provider | Context | Access |
|---|---|---|---|
| 01-ai/yi-large | NVIDIA NIM | 131K | Free tier |
| adept/fuyu-8b | NVIDIA NIM | 131K | Free tier |
| ai21labs/jamba-1.5-large-instruct | NVIDIA NIM | 131K | Free tier |
| aisingapore/sea-lion-7b-instruct | NVIDIA NIM | 131K | Free tier |
| baai/bge-m3 | NVIDIA NIM | 131K | Free tier |
Nemotron Mini 4B Instruct is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
Nemotron Mini 4B Instruct is listed with free API access on NVIDIA NIM, subject to the provider's quota and account policy.
The model ID shown in this catalog is nvidia/nemotron-mini-4b-instruct.
The listed context window is 128K tokens with up to 8K output tokens.
Free Nemotron Mini 4B Instruct API.
For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.