Cerebras logo

gemma-4-31b pricing and alternatives

Paid / limited
★★★★★★★★★★ 3.5 Benchmark-backed score

Largest Gemma 4 instruction model for open, self-hosted chat and reasoning

Paid / limitedOpenAI compatibleReasoningTool callingJSON modeTextImage
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.cerebras.ai/v1
Model ID
gemma-4-31b
API format OpenAI Chat Completions
No current free listing. This page is kept for comparison and SEO continuity. Check the alternatives below before integrating.
Technical Details

gemma-4-31b specifications

Provider catalog
Context window 131K
Max output 32K
Status Online
Family gemma
Released Apr 2, 2026
Last updated Aug 28, 2026
Free listing since Apr 2, 2026
Input text, image, reasoning
Output text
Capabilities reasoning, tool calling, structured output
Open weights Yes
AI Recommendation

Should you use gemma-4-31b?

gemma-4-31b is listed for chat workloads and supports a 131K context window.

Use the comparison and availability sections before choosing this listing, because it is not currently marked as free.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Structured JSON output

Watch outs

  • No current free listing on this provider
  • Free-tier rate limits apply
Benchmark Overview

Benchmark signals for gemma-4-31b

Measured data
Intelligence General reasoning and instruction following
29.4/100
Coding Programming and code generation
43.4/100
Agentic Tool use and multi-step tasks
14.4/100
Speed Observed generation speed
35 tok/s
Context Maximum listed context window
131K
Pricing

gemma-4-31b pricing per 1M tokens

Provider pricing
Input $0 per 1M tokens
Output $0 per 1M tokens
Free access Not listed Cerebras
Rate limit 15 RPM, 30K TPM, 1M TPD provider policy
Typical Use Cases

gemma-4-31b use cases

Chat

gemma-4-31b is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

gemma-4-31b free API FAQ

Is gemma-4-31b free to use?

gemma-4-31b is not currently marked as free on Cerebras.

What is the gemma-4-31b model ID?

The model ID shown in this catalog is gemma-4-31b.

What are the gemma-4-31b free tier rate limits on Cerebras?

The listed free tier limit is 15 RPM, 30K TPM, 1M TPD. Limits can change per account tier, so confirm against the provider dashboard.

What context window does gemma-4-31b support?

The listed context window is 131K tokens with up to 32K output tokens.

More about gemma-4-31b

Largest Gemma 4 instruction model for open, self-hosted chat and reasoning

For API keys, setup steps, and provider-level limits, see the Cerebras provider page.