Skip to main content

GPT-5.4 API

GPT-5.4 is a large language model by OpenAI, available on the Venice API as openai-gpt-54. Requests are anonymized, so the provider never sees your identity.GPT-5.4 is the latest frontier model in the GPT-5 series with a 1M+ context window, offering improved agentic and long context performance. It uses adaptive reasoning to dynamically allocate computation across tasks.

GPT-5.4 API pricing

GPT-5.4 specifications

How to use the GPT-5.4 API

Send requests to POST https://api.venice.ai/api/v1/responses with "model": "openai-gpt-54" and your API key.

GPT-5.4 API FAQ

How much does the GPT-5.4 API cost?

3.13per1Minputtokensand3.13 per 1M input tokens and 18.80 per 1M output tokens, with cached input at $0.31 per 1M. Prices are in USD and can be paid in DIEM at parity.

What is the GPT-5.4 model ID?

Use openai-gpt-54 as the model parameter.

Is the GPT-5.4 API private?

GPT-5.4 is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.

What is the context window of GPT-5.4?

1M tokens of context, with up to 128K output tokens per response.

What does GPT-5.4 support?

GPT-5.4 supports function calling, structured outputs, reasoning, image input, web search and prompt caching. Reasoning effort is adjustable with reasoning_effort: none, low, medium, high and xhigh (default high).

Which endpoint does the GPT-5.4 API use?

Call POST /responses. /chat/completions is also supported. OpenAI reasoning models are designed around the Responses API: reasoning, tool calls and messages come back as typed output items. Venice’s /responses endpoint is in alpha.

Related models