GPT-Realtime-2.1
Realtime API only. Successor to gpt-realtime-2 (2026-05-07, same prices, see gpt-realtime-2.md). Mini variant gpt-realtime-2.1-mini (audio $10/$20, text $0.60/$2.40). Replaces gpt-realtime / gpt-4o-realtime (shutdown Jan 20 2027). Azure version 2026-07-07.
- Context window
- 128,000 tokens
- Max output
- 32,000 tokens
- Knowledge cutoff
- 2024-09
- Input
- text, audio, image
- Output
- text, audio
- License
- proprietary
- Pricing
- text input: $4 · text output: $24 · audio input: $32 · audio output: $64 · cached input: $0.4 · image input: $5 (per 1M tokens (USD)) source
- Verified
- 2026-09-29
How to call it
| Provider | Model id | Endpoint / URL | Docs |
|---|---|---|---|
| OpenAI API | gpt-realtime-2.1 | wss://api.openai.com/v1/realtime | docs |
| Azure OpenAI (Microsoft Foundry) | gpt-realtime-2.1 | — | docs |
Notable capabilities (2)
- Reasoning in realtime voice: Configurable reasoning effort in a speech-to-speech model (at a latency cost). source
- Robust turn-taking: Improved alphanumeric recognition, silence/noise handling and interruption behavior. source
Low-latency voice agents over WebRTC/WebSocket (/v1/realtime?model=gpt-realtime-2.1).
Sources:
Other OpenAI models
GPT-6.1 Sol · GPT-6 Luna · GPT-6 Sol · GPT Image 2.5 Flare · GPT Image 2.5 Sunburst · GPT-6 Astra · GPT-Live-Transcribe · GPT-Transcribe · GPT-5.6 Terra · GPT-Live 1 · GPT-Realtime-2 · GPT-Realtime-Translate · GPT-Rosalind · GPT-Audio-1.5 (and gpt-audio / gpt-audio-mini) · GPT-5.3-Codex · gpt-oss-120b · gpt-oss-20b · GPT-4o mini TTS · text-embedding-3-large · text-embedding-3-small · GPT-5.6 Luna · GPT-5.6 Sol · GPT-Realtime-Whisper · GPT-5.5 Pro · GPT-5.5 · GPT Image 2 · GPT-5.4 · GPT-Realtime-1.5 · GPT-4.1 · GPT-4o · TTS-1 / TTS-1 HD · Whisper large-v3 / large-v3-turbo (open weights) · GPT-Realtime and GPT-Realtime mini · o3 · GPT-4o Transcribe / Mini Transcribe / Transcribe Diarize · Whisper (whisper-1 API) · Sora 2