NVIDIA NIM logo

Inkling API status on NVIDIA NIM

Check provider
Catalog metadata matched Reasoning models.dev metadata

Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio

Check providerOpenAI compatibleReasoningTool callingVisionAudioReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://integrate.api.nvidia.com/v1
Model ID
thinkingmachines/inkling
API format OpenAI Chat Completions
Technical Details

Inkling specifications

Provider and model catalog
Context window 256K
Max output 256K
Status Check provider
Family ling
Released Jul 15, 2026
Last updated Jul 15, 2026
Free listing since Jul 19, 2026
Input text, image, audio
Output text
Capabilities reasoning, tool calling, file attachments, temperature control
Open weights Yes
AI Recommendation

Should you use Inkling?

Inkling is listed for chat workloads and supports a 256K context window.

Check NVIDIA NIM's current endpoint status before integrating this listing. Use the alternatives table if the provider listing is unavailable.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Open weights available

Watch outs

  • Current endpoint availability needs provider confirmation
Pricing

Inkling pricing per 1M tokens

Check provider
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Check provider NVIDIA NIM
Availability

Inkling availability by provider

Current provider only

We only found this Inkling listing on NVIDIA NIM in the current catalog. Use the related models section for nearby alternatives.

Provider Model listing Access Context API Limits
NVIDIA NIM Inkling Check provider 256K Native Varies
View NVIDIA NIM setup guide →
Typical Use Cases

Inkling use cases

Chat

Inkling is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

Inkling free API FAQ

Is Inkling free to use?

Inkling appears in the free model catalog for NVIDIA NIM, but its current endpoint availability should be confirmed with the provider before use.

What is the Inkling model ID?

The model ID shown in this catalog is thinkingmachines/inkling.

What context window does Inkling support?

The listed context window is 256K tokens with up to 256K output tokens.

More about Inkling

Free Inkling API.

For API keys, setup steps, and provider-level limits, see the NVIDIA NIM provider page.