gemini-3.5-flash-lite · Google
Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for high-throughput, low-cost execution for subagent tasks and document parsing. The model supports text, image, video, audio, and PDF inputs, and is designed for high-volume agentic workflows, simple data extraction, and applications where latency and API cost are the primary constraints.
Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for high-throughput, low-cost execution for subagent tasks and document parsing. The model supports text, image, video, audio, and PDF inputs, and is designed for high-volume agentic workflows, simple data extraction, and applications where latency and API cost are the primary constraints.
Gemini 3.5 Flash Lite has a 1,048,576 token context window. It supports up to 65,536 output tokens.
On AIHubMix, Gemini 3.5 Flash Lite costs $0.3 per million input tokens and $2.5 per million output tokens. Cached input reads are billed at $0.03 per million tokens.
Gemini 3.5 Flash Lite accepts text, image, video, audio and PDF input.
Gemini 3.5 Flash Lite supports thinking, tool calling, function calling, structured outputs, web search and long context. Per-protocol parameter support is listed in the capability table on this page.
Gemini 3.5 Flash Lite is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemini-3.5-flash-lite — no other code changes needed.
Gemini 3.5 Flash Lite is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Gemini 3.5 Flash Lite was released on July 21, 2026 by Google.
Free version: Gemini 3.5 Flash Lite (free)
Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…
Gemini 3.7 Flash free version: Free model resources are limited and provided only for…
Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…
Google's newest, most compact, and most cost-effective image generation and editing…
Gemini 3.5 Flash-Lite free version: Free model resources are limited and provided only…
Gemini 3.6 Flash free version: Free model resources are limited and provided only for…
Use Gemini 3.5 Flash Lite via the AIHubMix unified API — one interface for every major LLM.