gpt-realtime-2.1 · OpenAI
GPT-Realtime-2.1 is a reasoning speech-to-speech model for the Realtime API, with tool use and configurable reasoning effort, improving on GPT-Realtime-2 in alphanumeric recognition, silence and noise handling, and interruption behavior.
GPT-Realtime-2.1 is a reasoning speech-to-speech model for the Realtime API, with tool use and configurable reasoning effort, improving on GPT-Realtime-2 in alphanumeric recognition, silence and noise handling, and interruption behavior.
GPT Realtime 2.1 has a 128,000 token context window.
On AIHubMix, GPT Realtime 2.1 costs $4 per million input tokens and $24 per million output tokens. Cached input reads are billed at $0.4 per million tokens.
GPT Realtime 2.1 accepts text, audio and image input.
GPT Realtime 2.1 is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gpt-realtime-2.1 — no other code changes needed.
GPT Realtime 2.1 is developed by OpenAI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
GPT Realtime 2.1 was released on July 6, 2026 by OpenAI.
GPT-6.1 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for…
GPT-6 Luna is OpenAI's latest and most efficient model, designed for focused…
GPT-6 Sol is OpenAI's latest model in the GPT-6 series, specifically designed for complex…
GPT-6 Astra is OpenAI's newest and most intelligent model, with industry-leading…
OpenAI's latest realtime speech-to-text model, built for low-latency use — it streams…
GPT-Image-2.5 Flare is OpenAI's latest image model, the fastest and suited for everyday…
Use GPT Realtime 2.1 via the AIHubMix unified API — one interface for every major LLM.