🎙 New speech recognition models from OpenAI

The company introduced GPT-transcribe for accurate audio file processing ($0.0045/min) and GPT-live-transcribe with low latency for live streaming ($0.017/min). Both models support multilingualism and contextual prompting.

🌍 Developers can now flexibly choose between archiving accuracy and the speed of interactive interfaces, optimizing costs and latency.

👤 This will enable the creation of cheaper and higher-quality tools for transcribing meetings, podcasts, and voice assistants.

Source 1: https://developers.openai.com/api/docs/models/gpt-transcribe Source 2: https://developers.openai.com/api/docs/models/gpt-live-transcribe