OpenAI launches next-generation audio models including GPT-4o Mini Transcribe
OpenAI announced gpt-4o-transcribe, gpt-4o-mini-transcribe, and gpt-4o-mini-tts β a new family of speech-to-text and text-to-speech models. The transcribe models employ RL-heavy training to achieve state-of-the-art WER, outperforming Whisper v2/v3 across FLEURS multilingual benchmarks.