Should you use glm-5.2?
glm-5.2 is listed for chat workloads and supports a 1.0M context window.
Check NVIDIA NIM's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.
https://integrate.api.nvidia.com/v1 z-ai/glm-5.2 glm-5.2 is listed for chat workloads and supports a 1.0M context window.
Check NVIDIA NIM's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.
We only found this glm-5.2 listing on NVIDIA NIM in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | z-ai/glm-5.2 | Check provider | 1.0M | OpenAI-style | Up to 40 RPM |
| Model | Provider | Context | Access |
|---|---|---|---|
| 01-ai/yi-large | NVIDIA NIM | 131K | Free tier |
| adept/fuyu-8b | NVIDIA NIM | 131K | Free tier |
| ai21labs/jamba-1.5-large-instruct | NVIDIA NIM | 131K | Free tier |
| aisingapore/sea-lion-7b-instruct | NVIDIA NIM | 131K | Free tier |
| baai/bge-m3 | NVIDIA NIM | 131K | Check provider |
glm-5.2 is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
glm-5.2 appears in the free model catalog for NVIDIA NIM, but its current endpoint availability should be confirmed with the provider before use.
The model ID shown in this catalog is z-ai/glm-5.2.
The listed free tier limit is Up to 40 RPM. Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 1.0M tokens with up to 262K output tokens.
Open flagship GLM for long-horizon coding agents and million-token context work
For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.