🎙 New speech recognition models from OpenAI
The company introduced GPT-transcribe for accurate audio file processing ($0.0045/min) and GPT-live-transcribe with low latency for live streaming ($0.017/min). Both models support multilingualism and contextual prompting.
🌍 Developers can now flexibly choose between archiving accuracy and the speed of interactive interfaces, optimizing costs and latency.
👤 This will enable the creation of cheaper and higher-quality tools for transcribing meetings, podcasts, and voice assistants.
Source 1: https://developers.openai.com/api/docs/models/gpt-transcribe Source 2: https://developers.openai.com/api/docs/models/gpt-live-transcribe
