GPT Realtime

Legacy OpenAI speech-to-speech realtime model, superseded by gpt-realtime-2.1. Served via the gateway's /v1/realtime WebSocket endpoint with text and audio input/output and function calling.

gpt-realtime
BETAScheduled for DeactivationGet StartedView uptime
32,768 context
Starting at $4.00/M input tokens
Starting at $16.00/M output tokens
Tools

Select Provider

All Providers for GPT Realtime

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

OpenAI
Context: 32.8k
Deprecated since Jul 20, 2026Deactivating on Jan 20, 2027
Input
$4
/M tokens
Cached
$0.4
/M tokens
Output
$16
/M tokens
Get Started