gemma-4-26b-a4b-it · Google
A Mixture-of-Experts model that activates only 4B parameters per inference,delivering high-performance reasoning with a fraction of the memory cost - idealfor cost-efficient, high-throughput server deployments.
A Mixture-of-Experts model that activates only 4B parameters per inference,delivering high-performance reasoning with a fraction of the memory cost - idealfor cost-efficient, high-throughput server deployments.
Gemma 4 26B A4B It has a 262,144 token context window.
On AIHubMix, Gemma 4 26B A4B It costs $0.14 per million input tokens and $0.4 per million output tokens.
Gemma 4 26B A4B It accepts text, image and video input.
Gemma 4 26B A4B It is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemma-4-26b-a4b-it — no other code changes needed.
Gemma 4 26B A4B It is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Gemma 4 26B A4B It was released on April 2, 2026 by Google.
Free version: Gemma 4 26B A4B It (free)
Gemini Nano Banana 2.1 (gemini-nano-banana-2.1) is Google's latest high-efficiency image…
Gemini 3.8 Flash-Lite TTS (gemini-3.8-flash-lite-tts) is Google's fast and affordable…
Gemini 3.8 Flash TTS (gemini-3.8-flash-tts) is Google's 3.8 Flash text-to-speech audio…
Gemini 3.8 Flash is Google's most intelligent Flash-series model, designed for…
Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…
Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…
Use Gemma 4 26B A4B It via the AIHubMix unified API — one interface for every major LLM.