OpenAI launches improved AI models for voice and transcription
OpenAI has introduced three new AI models designed to enhance speech-to-text and text-to-speech capabilities. The models gpt-4o-transcribe, gpt-4o-mini-transcribe, and gpt-4o-mini-tts offer improved accuracy and customization options for developers building voice applications. According to OpenAI, the new transcription models significantly outperform their predecessor, Whisper, particularly in noisy environments and with various accents. The company’s internal benchmarks …