Gemma 4 26B A4B It

gemma-4-26b-a4b-it · Google

A Mixture-of-Experts model that activates only 4B parameters per inference,delivering high-performance reasoning with a fraction of the memory cost - idealfor cost-efficient, high-throughput server deployments.

API Pricing

Input$0.14 / 1M tokens
Output$0.4 / 1M tokens

Specifications

Context262K tokens

Frequently asked questions

What is Gemma 4 26B A4B It?

A Mixture-of-Experts model that activates only 4B parameters per inference,delivering high-performance reasoning with a fraction of the memory cost - idealfor cost-efficient, high-throughput server deployments.

What is the context length of Gemma 4 26B A4B It?

Gemma 4 26B A4B It has a 262,100 token context window.

How much does Gemma 4 26B A4B It cost?

On AIHubMix, Gemma 4 26B A4B It costs $0.14 per million input tokens and $0.4 per million output tokens.

How do I call Gemma 4 26B A4B It via API?

Gemma 4 26B A4B It is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemma-4-26b-a4b-it — no other code changes needed.

Who created Gemma 4 26B A4B It?

Gemma 4 26B A4B It is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

Free version: Gemma 4 26B A4B It (free)

More models from Google

See all Google models →

Gemini 3.7 Flash

by Google

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…

$0.75/1M in · $3.75/1M out
1,048,576 tokens context

Gemini 3.7 Flash (free)

by Google

Gemini 3.7 Flash free version: Free model resources are limited and provided only for…

1,000,000 tokens context

Gemini 3.6 Flash

by Google

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $7.5/1M out
1,048,576 tokens context

Gemini 3.1 Flash Lite Image

by Google

Google's newest, most compact, and most cost-effective image generation and editing…

$0.25/1M in · $1.5/1M out

Gemini 3.5 Flash Lite

by Google

Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for…

$0.3/1M in · $2.5/1M out
1,048,576 tokens context

Gemini 3.5 Flash Lite (free)

by Google

Gemini 3.5 Flash-Lite free version: Free model resources are limited and provided only…

1,048,576 tokens context

Use Gemma 4 26B A4B It via the AIHubMix unified API — one interface for every major LLM.