o4-mini · OpenAI
o4-mini is a remarkably smart model for its speed and cost-efficiency. This allows it to support significantly higher usage limits than o3, making it a strong high-volume, high-throughput option for everyone with questions that benefit from reasoning.
o4-mini is a remarkably smart model for its speed and cost-efficiency. This allows it to support significantly higher usage limits than o3, making it a strong high-volume, high-throughput option for everyone with questions that benefit from reasoning.
O4 Mini has a 200,000 token context window.
On AIHubMix, O4 Mini costs $1.1 per million input tokens and $4.4 per million output tokens. Cached input reads are billed at $0.275 per million tokens.
O4 Mini accepts text and image input.
O4 Mini supports thinking, tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.
O4 Mini is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to o4-mini — no other code changes needed.
O4 Mini is developed by OpenAI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
GPT-5.6 Sol (limited-time 50% off) is OpenAI’s frontier reasoning model for complex…
GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly…
GPT‑5.6 Sol sets a new standard for both intelligence and efficiency, achieving…
GPT-5.6 Terra is designed for workloads that balance intelligence and cost. It roughly…
GPT-4o Transcribe Diarize is an automatic speech recognition (ASR) model with built-in…
The gpt-audio model is OpenAI's first officially released (generally available) audio…
Use O4 Mini via the AIHubMix unified API — one interface for every major LLM.