gemma-3n-e4b-it · Google
Gemma 3n is a generative AI model optimized for use in everyday devices, such as phones, laptops, and tablets. This model includes innovations in parameter-efficient processing, including Per-Layer Embedding (PLE) parameter caching and a MatFormer model architecture that provides the flexibility to reduce compute and memory requirements. These models feature audio input handling, as well as text and visual data.
Gemma 3n is a generative AI model optimized for use in everyday devices, such as phones, laptops, and tablets. This model includes innovations in parameter-efficient processing, including Per-Layer Embedding (PLE) parameter caching and a MatFormer model architecture that provides the flexibility to reduce compute and memory requirements. These models feature audio input handling, as well as text and visual data.
On AIHubMix, Gemma 3n E4B It costs $0.2 per million input tokens and $0.2 per million output tokens.
Gemma 3n E4B It is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemma-3n-e4b-it — no other code changes needed.
Gemma 3n E4B It is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Gemma 3n E4B It was released on May 20, 2025 by Google.
Gemini Nano Banana 2.1 (gemini-nano-banana-2.1) is Google's latest high-efficiency image…
Gemini 3.8 Flash-Lite TTS (gemini-3.8-flash-lite-tts) is Google's fast and affordable…
Gemini 3.8 Flash TTS (gemini-3.8-flash-tts) is Google's 3.8 Flash text-to-speech audio…
Gemini 3.8 Flash is Google's most intelligent Flash-series model, designed for…
Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…
Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…
Use Gemma 3n E4B It via the AIHubMix unified API — one interface for every major LLM.