Gemini 3.5 Flash

gemini-3.5-flash · Google

Gemini 3.5 Flash provides sustained frontier-level intelligence optimized for real-world tasks at a higher speed and lower cost. Designed for the agentic era, it excels at sub-agent deployment, multi-step workflows, and long-horizon tasks at scale. This model is particularly effective for rapid agentic loops involving complex coding cycles and iterations.

API Pricing

Input$1.5 / 1M tokens
Output$9 / 1M tokens
Cache read$0.15 / 1M tokens

Specifications

Context1.05M tokens
Modalitiestext, image, video, audio, PDF
CapabilitiesThinking, Tool calling, Web search, URL context, Code interpreter, Computer use, File search, Structured outputs, Prompt caching
Endpointschat_completions, gemini_api, claude_api

Frequently asked questions

What is Gemini 3.5 Flash?

Gemini 3.5 Flash provides sustained frontier-level intelligence optimized for real-world tasks at a higher speed and lower cost. Designed for the agentic era, it excels at sub-agent deployment, multi-step workflows, and long-horizon tasks at scale. This model is particularly effective for rapid agentic loops involving complex coding cycles and iterations.

What is the context length of Gemini 3.5 Flash?

Gemini 3.5 Flash has a 1,048,576 token context window.

How much does Gemini 3.5 Flash cost?

On AIHubMix, Gemini 3.5 Flash costs $1.5 per million input tokens and $9 per million output tokens. Cached input reads are billed at $0.15 per million tokens.

What modalities does Gemini 3.5 Flash support?

Gemini 3.5 Flash accepts text, image, video, audio and PDF input.

What capabilities does Gemini 3.5 Flash support?

Gemini 3.5 Flash supports thinking, tool calling, function calling, structured outputs, web search, deep search and long context. Per-protocol parameter support is listed in the capability table on this page.

How do I call Gemini 3.5 Flash via API?

Gemini 3.5 Flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemini-3.5-flash — no other code changes needed.

Who created Gemini 3.5 Flash?

Gemini 3.5 Flash is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was Gemini 3.5 Flash released?

Gemini 3.5 Flash was released on May 19, 2026 by Google.

More models from Google

See all Google models →

Gemini Nano Banana 2.1

by Google

Gemini Nano Banana 2.1 (gemini-nano-banana-2.1) is Google's latest high-efficiency image…

$1.5/1M in · $7.5/1M out
65,536 tokens context

Gemini 3.8 Flash Lite Tts

by Google

Gemini 3.8 Flash-Lite TTS (gemini-3.8-flash-lite-tts) is Google's fast and affordable…

$1/1M in · $1/1M out
8,192 tokens context

Gemini 3.8 Flash Tts

by Google

Gemini 3.8 Flash TTS (gemini-3.8-flash-tts) is Google's 3.8 Flash text-to-speech audio…

$0.5/1M in · $9/1M out
8,192 tokens context

Gemini 3.8 Flash

by Google

Gemini 3.8 Flash is Google's most intelligent Flash-series model, designed for…

$0.75/1M in · $3.75/1M out
1,048,576 tokens context

Gemini 3.7 Flash

by Google

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…

$0.75/1M in · $3.75/1M out
1,048,576 tokens context

Gemini 3.6 Flash

by Google

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $7.5/1M out
1,048,576 tokens context

Use Gemini 3.5 Flash via the AIHubMix unified API — one interface for every major LLM.