NVIDIA NIM logo

glm-5.2 API status on NVIDIA NIM

Check provider
★★★★★★★★★★ 3.5 Benchmark-backed score

Open flagship GLM for long-horizon coding agents and million-token context work

Check providerOpenAI compatibleReasoningTool callingJSON modeTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
z-ai/glm-5.2
API format OpenAI Chat Completions
Technical Details

glm-5.2 specifications

Provider catalog
Context window 1.0M
Max output 262K
Status Check provider
Released Jun 16, 2026
Last updated Aug 6, 2026
Free listing since Jun 16, 2026
Input text, reasoning
Output text
Capabilities reasoning, tool calling, structured output
AI Recommendation

Should you use glm-5.2?

glm-5.2 is listed for chat workloads and supports a 1.0M context window.

Check NVIDIA NIM's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Structured JSON output

Watch outs

  • Current endpoint availability needs provider confirmation
  • Free-tier rate limits apply
  • Vision support is not listed
Benchmark Overview

Benchmark signals for glm-5.2

Measured data
Intelligence General reasoning and instruction following
51.1/100
Coding Programming and code generation
68.8/100
Agentic Tool use and multi-step tasks
43.1/100
Speed Observed generation speed
148 tok/s
Context Maximum listed context window
1.0M
Pricing

glm-5.2 pricing per 1M tokens

Check provider
Input $1.35 per 1M tokens
Output $4.29 per 1M tokens
Free access Check provider NVIDIA NIM
Rate limit Up to 40 RPM provider policy
Availability

glm-5.2 availability by provider

Current provider only

We only found this glm-5.2 listing on NVIDIA NIM in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
NVIDIA NIM z-ai/glm-5.2 Check provider 1.0M OpenAI-style Up to 40 RPM
View NVIDIA NIM setup guide →
Typical Use Cases

glm-5.2 use cases

Chat

glm-5.2 is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

glm-5.2 free API FAQ

Is glm-5.2 free to use?

glm-5.2 appears in the free model catalog for NVIDIA NIM, but its current endpoint availability should be confirmed with the provider before use.

What is the glm-5.2 model ID?

The model ID shown in this catalog is z-ai/glm-5.2.

What are the glm-5.2 free tier rate limits on NVIDIA NIM?

The listed free tier limit is Up to 40 RPM. Limits can change per account tier, so confirm against the provider dashboard.

What context window does glm-5.2 support?

The listed context window is 1.0M tokens with up to 262K output tokens.

More about glm-5.2

Open flagship GLM for long-horizon coding agents and million-token context work

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.