o4-mini · OpenAI
o4-mini is a remarkably smart model for its speed and cost-efficiency. This allows it to support significantly higher usage limits than o3, making it a strong high-volume, high-throughput option for everyone with questions that benefit from reasoning.
o4-mini is a remarkably smart model for its speed and cost-efficiency. This allows it to support significantly higher usage limits than o3, making it a strong high-volume, high-throughput option for everyone with questions that benefit from reasoning.
O4 Mini has a 200,000 token context window.
On AIHubMix, O4 Mini costs $1.1 per million input tokens and $4.4 per million output tokens. Cached input reads are billed at $0.275 per million tokens.
O4 Mini accepts text, image and PDF input.
O4 Mini supports thinking, tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.
O4 Mini is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to o4-mini — no other code changes needed.
O4 Mini is developed by OpenAI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
O4 Mini was released on April 16, 2025 by OpenAI.
GPT-6.1 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for…
GPT-6 Luna is OpenAI's latest and most efficient model, designed for focused…
GPT-6 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for complex…
GPT-6 Astra is OpenAI's newest and most intelligent model, with industry-leading…
OpenAI's latest realtime speech-to-text model, built for low-latency use — it streams…
GPT-Realtime-2.1 is a reasoning speech-to-speech model for the Realtime API, with tool…
Use O4 Mini via the AIHubMix unified API — one interface for every major LLM.