GPT 4o

gpt-4o · OpenAI

GPT-4o (“o” stands for “omni”) is a new-generation multimodal model designed for more natural human–computer interaction. It can accept any combination of text, audio, image, and video as input, and generate multimodal outputs including text, audio, and images. With audio response latency as low as 232 milliseconds on average around 320 milliseconds, it approaches real human conversational speed. The model delivers strong performance in English text and code, significantly improved multilingual understanding, and outstanding capabilities in visual and audio perception, while offering faster API performance and substantially reduced cost for real-time and complex multimodal applications.

API Pricing

Input$2.5 / 1M tokens
Output$10 / 1M tokens
Cache read$1.25 / 1M tokens

Specifications

Context128K tokens
Modalitiestext, image, PDF
CapabilitiesStreaming, Tool calling, Web search, Code interpreter, File search, Structured outputs

Frequently asked questions

What is GPT 4o?

GPT-4o (“o” stands for “omni”) is a new-generation multimodal model designed for more natural human–computer interaction. It can accept any combination of text, audio, image, and video as input, and generate multimodal outputs including text, audio, and images. With audio response latency as low as 232 milliseconds on average around 320 milliseconds, it approaches real human conversational speed. The model delivers strong performance in English text and code, significantly improved multilingual understanding, and outstanding capabilities in visual and audio perception, while offering faster API performance and substantially reduced cost for real-time and complex multimodal applications.

What is the context length of GPT 4o?

GPT 4o has a 128,000 token context window.

How much does GPT 4o cost?

On AIHubMix, GPT 4o costs $2.5 per million input tokens and $10 per million output tokens. Cached input reads are billed at $1.25 per million tokens.

What modalities does GPT 4o support?

GPT 4o accepts text, image and PDF input.

What capabilities does GPT 4o support?

GPT 4o supports tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call GPT 4o via API?

GPT 4o is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gpt-4o — no other code changes needed.

Who created GPT 4o?

GPT 4o is developed by OpenAI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was GPT 4o released?

GPT 4o was released on May 13, 2024 by OpenAI.

Free version: GPT 4o (free)

More models from OpenAI

See all OpenAI models →

GPT 6.1 Sol

by OpenAI

GPT-6.1 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for…

$2/1M in · $10/1M out
1,050,000 tokens context

GPT 6 Luna

by OpenAI

GPT-6 Luna is OpenAI's latest and most efficient model, designed for focused…

$0.1/1M in · $0.5/1M out
1,050,000 tokens context

GPT 6 Sol

by OpenAI

GPT-6 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for complex…

$2/1M in · $10/1M out
1,050,000 tokens context

GPT 6 Astra

by OpenAI

GPT-6 Astra is OpenAI's newest and most intelligent model, with industry-leading…

$10/1M in · $50/1M out
1,050,000 tokens context

GPT Live Transcribe

by OpenAI

OpenAI's latest realtime speech-to-text model, built for low-latency use — it streams…

GPT Realtime 2.1

by OpenAI

GPT-Realtime-2.1 is a reasoning speech-to-speech model for the Realtime API, with tool…

$4/1M in · $24/1M out
128,000 tokens context

Use GPT 4o via the AIHubMix unified API — one interface for every major LLM.