Gemma 4 31B It

gemma-4-31b-it · Google

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function calling, and multilingual support across 140+ languages. Strong on coding, reasoning, and document understanding tasks. Apache 2.0 license.

API Pricing

Input$0.2 / 1M tokens
Output$0.4 / 1M tokens

Specifications

Context262K tokens
Modalitiestext, image, video
CapabilitiesThinking, Tool calling, Web search, Document input

Frequently asked questions

What is Gemma 4 31B It?

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function calling, and multilingual support across 140+ languages. Strong on coding, reasoning, and document understanding tasks. Apache 2.0 license.

What is the context length of Gemma 4 31B It?

Gemma 4 31B It has a 262,144 token context window.

How much does Gemma 4 31B It cost?

On AIHubMix, Gemma 4 31B It costs $0.2 per million input tokens and $0.4 per million output tokens.

What modalities does Gemma 4 31B It support?

Gemma 4 31B It accepts text, image and video input.

How do I call Gemma 4 31B It via API?

Gemma 4 31B It is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemma-4-31b-it — no other code changes needed.

Who created Gemma 4 31B It?

Gemma 4 31B It is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was Gemma 4 31B It released?

Gemma 4 31B It was released on April 2, 2026 by Google.

Free version: Gemma 4 31B It (free)

More models from Google

See all Google models →

Gemini Nano Banana 2.1

by Google

Gemini Nano Banana 2.1 (gemini-nano-banana-2.1) is Google's latest high-efficiency image…

$1.5/1M in · $7.5/1M out
65,536 tokens context

Gemini 3.8 Flash Lite Tts

by Google

Gemini 3.8 Flash-Lite TTS (gemini-3.8-flash-lite-tts) is Google's fast and affordable…

$1/1M in · $1/1M out
8,192 tokens context

Gemini 3.8 Flash Tts

by Google

Gemini 3.8 Flash TTS (gemini-3.8-flash-tts) is Google's 3.8 Flash text-to-speech audio…

$0.5/1M in · $9/1M out
8,192 tokens context

Gemini 3.8 Flash

by Google

Gemini 3.8 Flash is Google's most intelligent Flash-series model, designed for…

$0.75/1M in · $3.75/1M out
1,048,576 tokens context

Gemini 3.7 Flash

by Google

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…

$0.75/1M in · $3.75/1M out
1,048,576 tokens context

Gemini 3.6 Flash

by Google

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $7.5/1M out
1,048,576 tokens context

Use Gemma 4 31B It via the AIHubMix unified API — one interface for every major LLM.