Should you use llama-3.2-nemoretriever-1b-vlm-embed-v1?
llama-3.2-nemoretriever-1b-vlm-embed-v1 is listed for chat, embedding workloads and supports a 131K context window.
Use it when NVIDIA NIM's free tier is enough for evaluation, demos, or light production traffic.