gemini-2.5-flash · Google
Gemini 2.5 Flash is Google’s best model in terms of both performance and cost efficiency, offering a comprehensive set of capabilities. It is the first Flash model to support visible reasoning, allowing insight into the thought process behind its responses. With its strong price–performance ratio, the model is well suited for large-scale processing, low-latency, high-throughput tasks that require reasoning, as well as agent-based application scenarios.
Gemini 2.5 Flash is Google’s best model in terms of both performance and cost efficiency, offering a comprehensive set of capabilities. It is the first Flash model to support visible reasoning, allowing insight into the thought process behind its responses. With its strong price–performance ratio, the model is well suited for large-scale processing, low-latency, high-throughput tasks that require reasoning, as well as agent-based application scenarios.
Gemini 2.5 Flash has a 1,048,576 token context window.
On AIHubMix, Gemini 2.5 Flash costs $0.3 per million input tokens and $2.499 per million output tokens. Cached input reads are billed at $0.03 per million tokens.
Gemini 2.5 Flash accepts text, image, video, audio and PDF input.
Gemini 2.5 Flash supports tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.
Gemini 2.5 Flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemini-2.5-flash — no other code changes needed.
Gemini 2.5 Flash is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Gemini 2.5 Flash was released on June 17, 2025 by Google.
Gemini Nano Banana 2.1 (gemini-nano-banana-2.1) is Google's latest high-efficiency image…
Gemini 3.8 Flash-Lite TTS (gemini-3.8-flash-lite-tts) is Google's fast and affordable…
Gemini 3.8 Flash TTS (gemini-3.8-flash-tts) is Google's 3.8 Flash text-to-speech audio…
Gemini 3.8 Flash is Google's most intelligent Flash-series model, designed for…
Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…
Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…
Use Gemini 2.5 Flash via the AIHubMix unified API — one interface for every major LLM.