gpt-4o · OpenAI
GPT-4o (“o” stands for “omni”) is a new-generation multimodal model designed for more natural human–computer interaction. It can accept any combination of text, audio, image, and video as input, and generate multimodal outputs including text, audio, and images. With audio response latency as low as 232 milliseconds on average around 320 milliseconds, it approaches real human conversational speed. The model delivers strong performance in English text and code, significantly improved multilingual understanding, and outstanding capabilities in visual and audio perception, while offering faster API performance and substantially reduced cost for real-time and complex multimodal applications.
GPT-4o (“o” stands for “omni”) is a new-generation multimodal model designed for more natural human–computer interaction. It can accept any combination of text, audio, image, and video as input, and generate multimodal outputs including text, audio, and images. With audio response latency as low as 232 milliseconds on average around 320 milliseconds, it approaches real human conversational speed. The model delivers strong performance in English text and code, significantly improved multilingual understanding, and outstanding capabilities in visual and audio perception, while offering faster API performance and substantially reduced cost for real-time and complex multimodal applications.
GPT 4o has a 128,000 token context window.
On AIHubMix, GPT 4o costs $2.5 per million input tokens and $10 per million output tokens. Cached input reads are billed at $1.25 per million tokens.
GPT 4o accepts text, image and PDF input.
GPT 4o supports tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.
GPT 4o is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gpt-4o — no other code changes needed.
GPT 4o is developed by OpenAI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
GPT 4o was released on May 13, 2024 by OpenAI.
Free version: GPT 4o (free)
GPT-6.1 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for…
GPT-6 Luna is OpenAI's latest and most efficient model, designed for focused…
GPT-6 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for complex…
GPT-6 Astra is OpenAI's newest and most intelligent model, with industry-leading…
OpenAI's latest realtime speech-to-text model, built for low-latency use — it streams…
GPT-Realtime-2.1 is a reasoning speech-to-speech model for the Realtime API, with tool…
Use GPT 4o via the AIHubMix unified API — one interface for every major LLM.