Should you use llama-3.2-3b-instruct?
llama-3.2-3b-instruct is listed for chat workloads and supports a 131K context window.
Check NVIDIA NIM's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.