Gemma 3n E4B It

gemma-3n-e4b-it · Google

Gemma 3n is a generative AI model optimized for use in everyday devices, such as phones, laptops, and tablets. This model includes innovations in parameter-efficient processing, including Per-Layer Embedding (PLE) parameter caching and a MatFormer model architecture that provides the flexibility to reduce compute and memory requirements. These models feature audio input handling, as well as text and visual data.

API Pricing

Input$0.2 / 1M tokens
Output$0.2 / 1M tokens

Frequently asked questions

What is Gemma 3n E4B It?

Gemma 3n is a generative AI model optimized for use in everyday devices, such as phones, laptops, and tablets. This model includes innovations in parameter-efficient processing, including Per-Layer Embedding (PLE) parameter caching and a MatFormer model architecture that provides the flexibility to reduce compute and memory requirements. These models feature audio input handling, as well as text and visual data.

How much does Gemma 3n E4B It cost?

On AIHubMix, Gemma 3n E4B It costs $0.2 per million input tokens and $0.2 per million output tokens.

How do I call Gemma 3n E4B It via API?

Gemma 3n E4B It is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemma-3n-e4b-it — no other code changes needed.

Who created Gemma 3n E4B It?

Gemma 3n E4B It is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More models from Google

See all Google models →

Gemini 3.7 Flash

by Google

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…

$0.75/1M in · $3.75/1M out
1,048,576 tokens context

Gemini 3.7 Flash (free)

by Google

Gemini 3.7 Flash free version: Free model resources are limited and provided only for…

1,000,000 tokens context

Gemini 3.6 Flash

by Google

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $7.5/1M out
1,048,576 tokens context

Gemini 3.1 Flash Lite Image

by Google

Google's newest, most compact, and most cost-effective image generation and editing…

$0.25/1M in · $1.5/1M out

Gemini 3.5 Flash Lite

by Google

Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for…

$0.3/1M in · $2.5/1M out
1,048,576 tokens context

Gemini 3.5 Flash Lite (free)

by Google

Gemini 3.5 Flash-Lite free version: Free model resources are limited and provided only…

1,048,576 tokens context

Use Gemma 3n E4B It via the AIHubMix unified API — one interface for every major LLM.