Bullish
OpenAI Speech-to-Text Models Launch with 25% Price Cut and Lower Error Rates
2026-07-29 15:59:53
New GPT Transcribe models offer 25% lower costs and improved accuracy for batch and live audio, achieving a 3.31% word error rate.
Woofun AI reports that OpenAI has launched two new speech-to-text models, GPT Transcribe and GPT Live Transcribe, designed for batch processing and real-time captioning respectively. These models integrate audio topics and language cues to enhance recognition of technical terms, numbers, and accents in noisy environments.
Data from Artificial Analysis indicates GPT Transcribe achieves a 3.31% word error rate, a 0.7 percentage point improvement over the previous GPT-4o Transcribe. The pricing has been reduced by 25%, with costs set at $4.5 per 1000 minutes of audio.
WOOFUN AI
Impact Assessment · Quick Read
The reduction in both cost and error rates makes high-fidelity transcription more accessible for enterprise applications. Improved performance in noisy environments and with technical jargon could drive adoption in customer service and live broadcasting sectors. The price cut may intensify competition in the AI audio processing market.
Generated by WOOFUN AI · For reference only, not investment advice
Comments
No comments yet.