Gemini 2.5 Flash Preview 09 2025

gemini-2.5-flash-preview-09-2025 · Google

This latest 2.5 Flash model comes with improvements in two key areas we heard consistent feedback on: Better agentic tool use: We've improved how the model uses tools, leading to better performance in more complex, agentic and multi-step applications. This model shows noticeable improvements on key agentic benchmarks, including a 5% gain on SWE-Bench Verified, compared to our last release (48.9% → 54%). More efficient: With thinking on, the model is now significantly more cost-efficient—achieving higher quality outputs while using fewer tokens, reducing latency and cost (see charts above).

API Pricing

Input$0.3 / 1M tokens
Output$2.499 / 1M tokens
Cache read$0.03 / 1M tokens

Specifications

Context1.05M tokens
Modalitiestext, image, video, audio
CapabilitiesThinking, Tool calling, Web search, URL context, Code interpreter, File search, Structured outputs, Prompt caching

Frequently asked questions

What is Gemini 2.5 Flash Preview 09 2025?

This latest 2.5 Flash model comes with improvements in two key areas we heard consistent feedback on: Better agentic tool use: We've improved how the model uses tools, leading to better performance in more complex, agentic and multi-step applications. This model shows noticeable improvements on key agentic benchmarks, including a 5% gain on SWE-Bench Verified, compared to our last release (48.9% → 54%). More efficient: With thinking on, the model is now significantly more cost-efficient—achieving higher quality outputs while using fewer tokens, reducing latency and cost (see charts above).

What is the context length of Gemini 2.5 Flash Preview 09 2025?

Gemini 2.5 Flash Preview 09 2025 has a 1,048,576 token context window.

How much does Gemini 2.5 Flash Preview 09 2025 cost?

On AIHubMix, Gemini 2.5 Flash Preview 09 2025 costs $0.3 per million input tokens and $2.499 per million output tokens. Cached input reads are billed at $0.03 per million tokens.

What modalities does Gemini 2.5 Flash Preview 09 2025 support?

Gemini 2.5 Flash Preview 09 2025 accepts text, image, video and audio input.

What capabilities does Gemini 2.5 Flash Preview 09 2025 support?

Gemini 2.5 Flash Preview 09 2025 supports tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call Gemini 2.5 Flash Preview 09 2025 via API?

Gemini 2.5 Flash Preview 09 2025 is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemini-2.5-flash-preview-09-2025 — no other code changes needed.

Who created Gemini 2.5 Flash Preview 09 2025?

Gemini 2.5 Flash Preview 09 2025 is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was Gemini 2.5 Flash Preview 09 2025 released?

Gemini 2.5 Flash Preview 09 2025 was released on September 25, 2025 by Google.

More models from Google

See all Google models →

Gemini Nano Banana 2.1

by Google

Gemini Nano Banana 2.1 (gemini-nano-banana-2.1) is Google's latest high-efficiency image…

$1.5/1M in · $7.5/1M out
65,536 tokens context

Gemini 3.8 Flash Lite Tts

by Google

Gemini 3.8 Flash-Lite TTS (gemini-3.8-flash-lite-tts) is Google's fast and affordable…

$1/1M in · $1/1M out
8,192 tokens context

Gemini 3.8 Flash Tts

by Google

Gemini 3.8 Flash TTS (gemini-3.8-flash-tts) is Google's 3.8 Flash text-to-speech audio…

$0.5/1M in · $9/1M out
8,192 tokens context

Gemini 3.8 Flash

by Google

Gemini 3.8 Flash is Google's most intelligent Flash-series model, designed for…

$0.75/1M in · $3.75/1M out
1,048,576 tokens context

Gemini 3.7 Flash

by Google

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…

$0.75/1M in · $3.75/1M out
1,048,576 tokens context

Gemini 3.6 Flash

by Google

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…

$1.5/1M in · $7.5/1M out
1,048,576 tokens context

Use Gemini 2.5 Flash Preview 09 2025 via the AIHubMix unified API — one interface for every major LLM.