o3-mini · OpenAI
OpenAI's latest fast inference model excels at STEAM tasks and offers exceptional cost-effectiveness. Official support for cache hits reduces input prices by half.
OpenAI's latest fast inference model excels at STEAM tasks and offers exceptional cost-effectiveness. Official support for cache hits reduces input prices by half.
O3 Mini has a 200,000 token context window.
On AIHubMix, O3 Mini costs $1.1 per million input tokens and $4.4 per million output tokens. Cached input reads are billed at $0.55 per million tokens.
O3 Mini accepts text and image input.
O3 Mini supports thinking. Per-protocol parameter support is listed in the capability table on this page.
O3 Mini is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to o3-mini — no other code changes needed.
O3 Mini is developed by OpenAI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
O3 Mini was released on December 20, 2024 by OpenAI.
GPT-6.1 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for…
GPT-6 Luna is OpenAI's latest and most efficient model, designed for focused…
GPT-6 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for complex…
GPT-6 Astra is OpenAI's newest and most intelligent model, with industry-leading…
OpenAI's latest realtime speech-to-text model, built for low-latency use — it streams…
GPT-Realtime-2.1 is a reasoning speech-to-speech model for the Realtime API, with tool…
Use O3 Mini via the AIHubMix unified API — one interface for every major LLM.