gpt-5.4-low

by OpenAI

GPT-5.4 supports configuring reasoning strength only through the /responses endpoint. To make lower-overhead reasoning available directly in the /chat endpoint, the GPT-5.4-Low model is provided. This model is based on GPT-5.4 with reasoning_effort preset to low. This model is designed for use cases that are sensitive to response latency and cost. By adopting a lighter reasoning strategy, it delivers stable responses with lower latency and higher throughput. It is well suited for high-concurrency conversations, real-time interactions, basic Q&A, and scenarios where deep reasoning is not required.

API Pricing

Input$2.5 / 1M tokens
Output$15 / 1M tokens
Cache read$0.25 / 1M tokens

Specifications

Context window400,000 tokens
Modalitiestext, image
Featuresthinking, function calling, web, structured outputs, tool calling

FAQ

What is gpt-5.4-low?

GPT-5.4 supports configuring reasoning strength only through the /responses endpoint. To make lower-overhead reasoning available directly in the /chat endpoint, the GPT-5.4-Low model is provided. This model is based on GPT-5.4 with reasoning_effort preset to low. This model is designed for use cases that are sensitive to response latency and cost. By adopting a lighter reasoning strategy, it delivers stable responses with lower latency and higher throughput. It is well suited for high-concurrency conversations, real-time interactions, basic Q&A, and scenarios where deep reasoning is not required.

What is the context length of gpt-5.4-low?

gpt-5.4-low has a 400,000 token context window.

How much does gpt-5.4-low cost?

On AIHubMix, gpt-5.4-low costs $2.5 per million input tokens and $15 per million output tokens. Cached input reads are billed at $0.25 per million tokens.

What modalities does gpt-5.4-low support?

gpt-5.4-low accepts text and image input.

What features does gpt-5.4-low support?

gpt-5.4-low supports thinking, function calling, web, structured outputs and tool calling. Per-protocol parameter support is listed in the capability table on this page.

How do I call gpt-5.4-low via API?

gpt-5.4-low is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gpt-5.4-low — no other code changes needed.

Who develops gpt-5.4-low?

gpt-5.4-low is developed by OpenAI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More from OpenAI

gpt-5.6-luna

by OpenAI

GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly…

$0.2/1M in · $1.2/1M out
1,050,000 tokens context

gpt-5.6-sol

by OpenAI

GPT‑5.6 Sol sets a new standard for both intelligence and efficiency, achieving…

$5/1M in · $30/1M out
1,050,000 tokens context

gpt-5.6-terra

by OpenAI

GPT-5.6 Terra is designed for workloads that balance intelligence and cost. It roughly…

$2/1M in · $12/1M out
1,050,000 tokens context

gpt-4o-transcribe-diarize

by OpenAI

GPT-4o Transcribe Diarize is an automatic speech recognition (ASR) model with built-in…

$2.5/1M in · $10/1M out
16,000 tokens context

gpt-audio-1.5

by OpenAI

The gpt-audio model is OpenAI's first officially released (generally available) audio…

$2.5/1M in · $10/1M out
128,000 tokens context

gpt-image-2

by OpenAI

GPT-image-2 is OpenAI's latest cutting-edge image generation model. Key value adds…

$5/1M in · $30/1M out

Use gpt-5.4-low via the AIHubMix unified API — one interface for every major LLM.