# Advancing voice intelligence with new models in the API

DevFeed: [Advancing voice intelligence with new models in the API](<https://devfeed.tech/articles/advancing-voice-intelligence-with-new-models-in-the-api-6280.md>)

Original publisher: [Read original article](<https://openai.com/index/advancing-voice-intelligence-with-new-models-in-the-api>)

Published: 2026-05-07T10:00:00Z

Content type: release

Language: en

Sources: [OpenAI News](<https://devfeed.tech/sources/openai-news.md>)

Topics: [API](<https://devfeed.tech/topics/api.md>), [voice ai](<https://devfeed.tech/topics/voice-ai.md>), [OpenAI](<https://devfeed.tech/topics/openai.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [asr](<https://devfeed.tech/topics/asr.md>), [Streaming](<https://devfeed.tech/topics/streaming.md>), [Whisper](<https://devfeed.tech/topics/whisper.md>)

Tags: [agent](<https://devfeed.tech/tags/agent.md>), [api](<https://devfeed.tech/tags/api.md>), [audio](<https://devfeed.tech/tags/audio.md>), [developers](<https://devfeed.tech/tags/developers.md>), [launch](<https://devfeed.tech/tags/launch.md>), [model](<https://devfeed.tech/tags/model.md>), [models](<https://devfeed.tech/tags/models.md>), [openai](<https://devfeed.tech/tags/openai.md>), [product](<https://devfeed.tech/tags/product.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [speech](<https://devfeed.tech/tags/speech.md>), [streaming](<https://devfeed.tech/tags/streaming.md>), [tools](<https://devfeed.tech/tags/tools.md>), [voice](<https://devfeed.tech/tags/voice.md>), [voice-ai](<https://devfeed.tech/tags/voice-ai.md>), [whisper](<https://devfeed.tech/tags/whisper.md>)

## AI overview

OpenAI is introducing three audio models in its API: GPT-Realtime-2 for more capable conversational voice interactions, GPT-Realtime-Translate for live speech translation, and GPT-Realtime-Whisper for streaming speech-to-text. The models are designed to support voice applications that can listen, reason, translate, transcribe, use tools, and take action in real time.

## Source excerpt

Explore new realtime voice models in the OpenAI API that can reason, translate, and transcribe speech, enabling more natural and intelligent voice experiences.