Skip to content
Sovyron

GPT-Realtime-2.1

gpt128k context32k max outputreleased Jul 6, 2026ReasoningTool calling

As of Oct 4, 2026, the cheapest GPT-Realtime-2.1 API price is $4 per 1M input tokens (OpenAI) and $24 per 1M output tokens (OpenAI). GPT-Realtime-2.1 is offered by 2 providers in the gpt family, reasoning enabled.

Cheapest input
$4
Cheapest output
$24
Official (OpenAI)
$24
output per 1M tokens
Providers
2
offering this model
Spread
1.0×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

GPT-Realtime-2.1 API prices by provider in USD per 1M tokens, updated Oct 4, 2026
OpenAICheapest$4Cheapest$24$0.4—128k
Vercel AI Gateway$4$24$0.4—128k

GPT-Realtime-2.1 pricing FAQ

What is the cheapest GPT-Realtime-2.1 API?

As of Oct 4, 2026, OpenAI has the lowest GPT-Realtime-2.1 output price at $24 per 1M tokens, and OpenAI has the lowest input price at $4 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does GPT-Realtime-2.1 cost on OpenAI?

OpenAI charges $4 per 1M input tokens and $24 per 1M output tokens for GPT-Realtime-2.1.

How many providers offer GPT-Realtime-2.1?

2 providers list GPT-Realtime-2.1 on Sovyron; 2 of them sell it at a metered per-token price.

Is GPT-Realtime-2.1 free?

No provider in the Sovyron catalog lists a free tier for GPT-Realtime-2.1.

What is the context window of GPT-Realtime-2.1?

GPT-Realtime-2.1 supports a 128k-token context window and up to 32k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—