GPT 5.4 Low

gpt-5.4-low · OpenAI

GPT-5.4 supports configuring reasoning strength only through the /responses endpoint. To make lower-overhead reasoning available directly in the /chat endpoint, the GPT-5.4-Low model is provided. This model is based on GPT-5.4 with reasoning_effort preset to low. This model is designed for use cases that are sensitive to response latency and cost. By adopting a lighter reasoning strategy, it delivers stable responses with lower latency and higher throughput. It is well suited for high-concurrency conversations, real-time interactions, basic Q&A, and scenarios where deep reasoning is not required.

API Pricing

Input$2.5 / 1M tokens
Output$15 / 1M tokens
Cache read$0.25 / 1M tokens

Specifications

Context400K tokens
Modalitiestext, image
Featuresthinking, function calling, web search, structured outputs, tool calling

Frequently asked questions

What is GPT 5.4 Low?

GPT-5.4 supports configuring reasoning strength only through the /responses endpoint. To make lower-overhead reasoning available directly in the /chat endpoint, the GPT-5.4-Low model is provided. This model is based on GPT-5.4 with reasoning_effort preset to low. This model is designed for use cases that are sensitive to response latency and cost. By adopting a lighter reasoning strategy, it delivers stable responses with lower latency and higher throughput. It is well suited for high-concurrency conversations, real-time interactions, basic Q&A, and scenarios where deep reasoning is not required.

What is the context length of GPT 5.4 Low?

GPT 5.4 Low has a 400,000 token context window.

How much does GPT 5.4 Low cost?

On AIHubMix, GPT 5.4 Low costs $2.5 per million input tokens and $15 per million output tokens. Cached input reads are billed at $0.25 per million tokens.

What modalities does GPT 5.4 Low support?

GPT 5.4 Low accepts text and image input.

What capabilities does GPT 5.4 Low support?

GPT 5.4 Low supports thinking, function calling, web search, structured outputs and tool calling. Per-protocol parameter support is listed in the capability table on this page.

How do I call GPT 5.4 Low via API?

GPT 5.4 Low is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gpt-5.4-low — no other code changes needed.

Who created GPT 5.4 Low?

GPT 5.4 Low is developed by OpenAI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More models from OpenAI

See all OpenAI models →

GPT 5.6 Sol Disc

by OpenAI

GPT-5.6 Sol (limited-time 50% off) is OpenAI’s frontier reasoning model for complex…

$5 $2.5/1M in · $30 $15/1M out
50% off · 00:00–23:59 UTC
1,050,000 tokens context

GPT 5.6 Luna

by OpenAI

GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly…

$0.2/1M in · $1.2/1M out
1,050,000 tokens context

GPT 5.6 Sol

by OpenAI

GPT‑5.6 Sol sets a new standard for both intelligence and efficiency, achieving…

$5/1M in · $30/1M out
1,050,000 tokens context

GPT 5.6 Terra

by OpenAI

GPT-5.6 Terra is designed for workloads that balance intelligence and cost. It roughly…

$2/1M in · $12/1M out
1,050,000 tokens context

GPT 4o Transcribe Diarize

by OpenAI

GPT-4o Transcribe Diarize is an automatic speech recognition (ASR) model with built-in…

$2.5/1M in · $10/1M out
16,000 tokens context

GPT Audio 1.5

by OpenAI

The gpt-audio model is OpenAI's first officially released (generally available) audio…

$2.5/1M in · $10/1M out
128,000 tokens context

Use GPT 5.4 Low via the AIHubMix unified API — one interface for every major LLM.