GPT-Realtime-1.5
Released 2026-02-23 alongside gpt-audio-1.5 (Chat Completions). Docs still call it 'our flagship audio model for voice agents', but gpt-realtime-2 (May 2026) and gpt-realtime-2.1 (July 2026) supersede it; not deprecated as of 2026-09-29. It is the named replacement for the gpt-4o-realtime-preview models shut down 2026-05-12.
- Context window
- 32,000 tokens
- Max output
- 4,096 tokens
- Knowledge cutoff
- 2024-09
- Input
- text, audio, image
- Output
- text, audio
- License
- proprietary
- Pricing
- text input: $4 · text output: $16 · audio input: $32 · audio output: $64 · cached input: $0.4 · image input: $5 (per 1M tokens (USD)) source
- Verified
- 2026-09-29
How to call it
| Provider | Model id | Endpoint / URL | Docs |
|---|---|---|---|
| OpenAI API | gpt-realtime-1.5 | wss://api.openai.com/v1/realtime | docs |
Notable capabilities (1)
- Non-reasoning voice agent model: Speech-to-speech model for voice agents and customer support with function calling and prompt caching; cheaper text output ($16 vs $24/1M) than the reasoning gpt-realtime-2.x models. source
Sources:
Other OpenAI models
GPT-6.1 Sol · GPT-6 Luna · GPT-6 Sol · GPT Image 2.5 Flare · GPT Image 2.5 Sunburst · GPT-6 Astra · GPT-Live-Transcribe · GPT-Transcribe · GPT-5.6 Terra · GPT-Live 1 · GPT-Realtime-2.1 · GPT-Realtime-2 · GPT-Realtime-Translate · GPT-Rosalind · GPT-Audio-1.5 (and gpt-audio / gpt-audio-mini) · GPT-5.3-Codex · gpt-oss-120b · gpt-oss-20b · GPT-4o mini TTS · text-embedding-3-large · text-embedding-3-small · GPT-5.6 Luna · GPT-5.6 Sol · GPT-Realtime-Whisper · GPT-5.5 Pro · GPT-5.5 · GPT Image 2 · GPT-5.4 · GPT-4.1 · GPT-4o · TTS-1 / TTS-1 HD · Whisper large-v3 / large-v3-turbo (open weights) · GPT-Realtime and GPT-Realtime mini · o3 · GPT-4o Transcribe / Mini Transcribe / Transcribe Diarize · Whisper (whisper-1 API) · Sora 2