GPT Realtime 2.1

OpenAI's current speech-to-speech realtime model. Served via the gateway's /v1/realtime WebSocket endpoint with text and audio input/output and function calling.

gpt-realtime-2.1
BETAGet StartedView uptime
128,000 context
Starting at $4.00/M input tokens
Starting at $24.00/M output tokens
Tools
Reasoning

Select Provider

All Providers for GPT Realtime 2.1

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

OpenAI
Context: 128k
Input
$4
/M tokens
Cached
$0.4
/M tokens
Output
$24
/M tokens
Get Started