O4 Mini

o4-mini · OpenAI

o4-mini is a remarkably smart model for its speed and cost-efficiency. This allows it to support significantly higher usage limits than o3, making it a strong high-volume, high-throughput option for everyone with questions that benefit from reasoning.

API Pricing

Input$1.1 / 1M tokens
Output$4.4 / 1M tokens
Cache read$0.275 / 1M tokens

Specifications

Context200K tokens
Modalitiestext, image, PDF
CapabilitiesThinking, Streaming, Tool calling, Web search, Code interpreter, File search, Structured outputs, Prompt caching

Frequently asked questions

What is O4 Mini?

o4-mini is a remarkably smart model for its speed and cost-efficiency. This allows it to support significantly higher usage limits than o3, making it a strong high-volume, high-throughput option for everyone with questions that benefit from reasoning.

What is the context length of O4 Mini?

O4 Mini has a 200,000 token context window.

How much does O4 Mini cost?

On AIHubMix, O4 Mini costs $1.1 per million input tokens and $4.4 per million output tokens. Cached input reads are billed at $0.275 per million tokens.

What modalities does O4 Mini support?

O4 Mini accepts text, image and PDF input.

What capabilities does O4 Mini support?

O4 Mini supports thinking, tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call O4 Mini via API?

O4 Mini is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to o4-mini — no other code changes needed.

Who created O4 Mini?

O4 Mini is developed by OpenAI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was O4 Mini released?

O4 Mini was released on April 16, 2025 by OpenAI.

More models from OpenAI

See all OpenAI models →

GPT 6.1 Sol

by OpenAI

GPT-6.1 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for…

$2/1M in · $10/1M out
1,050,000 tokens context

GPT 6 Luna

by OpenAI

GPT-6 Luna is OpenAI's latest and most efficient model, designed for focused…

$0.1/1M in · $0.5/1M out
1,050,000 tokens context

GPT 6 Sol

by OpenAI

GPT-6 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for complex…

$2/1M in · $10/1M out
1,050,000 tokens context

GPT 6 Astra

by OpenAI

GPT-6 Astra is OpenAI's newest and most intelligent model, with industry-leading…

$10/1M in · $50/1M out
1,050,000 tokens context

GPT Live Transcribe

by OpenAI

OpenAI's latest realtime speech-to-text model, built for low-latency use — it streams…

GPT Realtime 2.1

by OpenAI

GPT-Realtime-2.1 is a reasoning speech-to-speech model for the Realtime API, with tool…

$4/1M in · $24/1M out
128,000 tokens context

Use O4 Mini via the AIHubMix unified API — one interface for every major LLM.