Gemini 3.1 Flash Lite

gemini-3.1-flash-lite · Google

gemini-3.1-flash-lite is currently Google's latest and most cost-effective model, optimized for large-scale agent-based tasks, translation, and simple data processing.

API Pricing

Input$0.25 / 1M tokens
Output$1.5 / 1M tokens

Specifications

Context1M tokens
Modalitiestext, image, audio, video
Featuresthinking, tool calling, function calling, structured outputs, web search, deep search, long context
Endpointschat_completions, gemini_api, claude_api

Frequently asked questions

What is Gemini 3.1 Flash Lite?

gemini-3.1-flash-lite is currently Google's latest and most cost-effective model, optimized for large-scale agent-based tasks, translation, and simple data processing.

What is the context length of Gemini 3.1 Flash Lite?

Gemini 3.1 Flash Lite has a 1,000,000 token context window.

How much does Gemini 3.1 Flash Lite cost?

On AIHubMix, Gemini 3.1 Flash Lite costs $0.25 per million input tokens and $1.5 per million output tokens.

What modalities does Gemini 3.1 Flash Lite support?

Gemini 3.1 Flash Lite accepts text, image, audio and video input.

What capabilities does Gemini 3.1 Flash Lite support?

Gemini 3.1 Flash Lite supports thinking, tool calling, function calling, structured outputs, web search, deep search and long context. Per-protocol parameter support is listed in the capability table on this page.

How do I call Gemini 3.1 Flash Lite via API?

Gemini 3.1 Flash Lite is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemini-3.1-flash-lite — no other code changes needed.

Who created Gemini 3.1 Flash Lite?

Gemini 3.1 Flash Lite is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More models from Google

See all Google models →

Gemini 3.7 Flash

by Google

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…

$0.75/1M in · $3.75/1M out
1,048,576 tokens context

Gemini 3.7 Flash (free)

by Google

Gemini 3.7 Flash free version: Free model resources are limited and provided only for…

1,000,000 tokens context

Gemini 3.6 Flash

by Google

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $7.5/1M out
1,048,576 tokens context

Gemini 3.1 Flash Lite Image

by Google

Google's newest, most compact, and most cost-effective image generation and editing…

$0.25/1M in · $1.5/1M out

Gemini 3.5 Flash Lite

by Google

Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for…

$0.3/1M in · $2.5/1M out
1,048,576 tokens context

Gemini 3.5 Flash Lite (free)

by Google

Gemini 3.5 Flash-Lite free version: Free model resources are limited and provided only…

1,048,576 tokens context

Use Gemini 3.1 Flash Lite via the AIHubMix unified API — one interface for every major LLM.