🗣️ OpenAI Unveils Next-Gen Voice AI

OpenAI has launched a powerful new lineup of audio AI models aimed at revolutionizing voice technology. The release includes gpt-4o-mini-tts, a text-to-speech model with fine-tuned control over tone and timing, and two advanced speech-to-text models, gpt-4o-transcribe and gpt-4o-mini-transcribe, which outperform the previous Whisper model, especially in noisy environments and with diverse accents. Available via OpenAI’s API and Agents SDK, these models are competitively priced, enabling developers to build more expressive voice applications and accurate transcription tools. OpenAI also launched OpenAI FM, a platform to demo the text-to-speech tech, along with a contest to spur creative uses. The release has sparked major interest across the tech community, signaling a new era for voice-driven AI.
