Should you use llama-nemotron-embed-1b-v2?
llama-nemotron-embed-1b-v2 is listed for chat, reasoning, embedding workloads and supports a 131K context window.
Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.