Article
OpenAI introduces new speech transcription models
The new models power both batch processing and real-time streaming for speech-to-text tasks.
29 Jul 20261 min readAI4U Desk

Artificial intelligence is making communication easier than ever with the introduction of gpt-transcribe and gpt-live-transcribe. These new tools are designed to convert speech into text seamlessly, opening up exciting possibilities for content creators, students, and professionals alike.
The models power both batch processing and real-time streaming. Users can prompt them with specific domain terms to ensure accurate spelling of rare words using custom vocabulary features. Additionally, the models filter out heavy background interference while accurately translating diverse languages.
For offline analysis and record keeping, gpt-transcribe efficiently handles large audio files through batch processing. Meanwhile, gpt-live-transcribe unlocks instant speech-to-text use cases for real-time applications.


