NVIDIA NIM logo

llama-3.2-1b-instruct API status on NVIDIA NIM

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

Llama 3.2 1B Instruct is Meta's smallest instruction-tuned model, available free on NVIDIA NIM.

Check providerOpenAI compatibleText
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
meta/llama-3.2-1b-instruct
API format OpenAI Chat Completions
Technical Details

llama-3.2-1b-instruct specifications

Provider catalog
Context window 60K
Max output 60K
Status Check provider
Family llama-3-2-1b-instruct
Knowledge cutoff 2023-12
Released Sep 25, 2024
Last updated Aug 6, 2026
Free listing since Sep 25, 2024
Input text
Output text
Open weights Yes
AI Recommendation

Should you use llama-3.2-1b-instruct?

llama-3.2-1b-instruct is listed for chat workloads and supports a 60K context window.

Check NVIDIA NIM's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Works with OpenAI-style SDKs

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
  • Tool calling is not confirmed
Benchmark Overview

Benchmark signals for llama-3.2-1b-instruct

Measured data
Intelligence General reasoning and instruction following
6.3/100
Coding Programming and code generation
0.6/100
Context Maximum listed context window
60K
Pricing

llama-3.2-1b-instruct pricing per 1M tokens

Check provider
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Check provider NVIDIA NIM
Rate limit Up to 40 RPM provider policy
Typical Use Cases

llama-3.2-1b-instruct use cases

Chat

llama-3.2-1b-instruct is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

llama-3.2-1b-instruct free API FAQ

Is llama-3.2-1b-instruct free to use?

llama-3.2-1b-instruct appears in the free model catalog for NVIDIA NIM, but its current endpoint availability should be confirmed with the provider before use.

What is the llama-3.2-1b-instruct model ID?

The model ID shown in this catalog is meta/llama-3.2-1b-instruct.

What are the llama-3.2-1b-instruct free tier rate limits on NVIDIA NIM?

The listed free tier limit is Up to 40 RPM. Limits can change per account tier, so confirm against the provider dashboard.

What context window does llama-3.2-1b-instruct support?

The listed context window is 60K tokens with up to 60K output tokens.

More about llama-3.2-1b-instruct

Llama 3.2 1B Instruct is Meta's smallest instruction-tuned model, available free on NVIDIA NIM. At 1B parameters, it is extremely fast and lightweight — suitable for high-throughput classification, simple Q&A, and edge-case testing. Up to 40 RPM, no daily token cap. NVIDIA's OpenAI-compatible API; requires free Developer Program membership and phone verification.

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.