Modelwire
Subscribe

OpenAI launches speech models but lags rivals on accuracy

Illustration accompanying: GPT Transcribe improves on its predecessor but can't catch ElevenLabs, Google, or Mistral on error rates

OpenAI has entered the speech recognition market with GPT Transcribe and GPT Live Transcribe, API-accessible models that represent incremental progress over prior versions but trail established competitors on error rates. The release signals OpenAI's push into multimodal infrastructure, though the positioning as a follower rather than leader suggests the company is playing catch-up in a crowded segment dominated by ElevenLabs, Google, and Mistral. For developers, this expands API options but doesn't reshape the competitive landscape; the real question is whether OpenAI can iterate toward parity or if speech recognition remains a secondary priority.

Modelwire context

Skeptical read

OpenAI hasn't disclosed the actual error rate deltas versus its predecessor, only that GPT Transcribe improved and still underperforms rivals. The absence of head-to-head benchmarks on identical test sets (vs. vendor-reported numbers) is the qualifier worth noting.

This is largely disconnected from recent activity in the space. We have no prior Modelwire coverage of OpenAI's speech recognition efforts or the competitive positioning of ElevenLabs, Google, and Mistral in transcription. This story belongs to the broader pattern of OpenAI expanding into infrastructure layers (APIs, modalities) where it doesn't yet lead, a dynamic we should track as it repeats across new domains.

If OpenAI publishes independent third-party benchmarks (not self-reported) showing GPT Transcribe within 5% of ElevenLabs' error rate within the next two quarters, that signals serious investment. If the company ships no new benchmark data and focuses instead on pricing or latency claims, treat this as a feature release, not a competitive entry.

This analysis is generated by Modelwire’s editorial layer from our archive and the summary above. It is not a substitute for the original reporting. How we write it.

MentionsOpenAI · GPT Transcribe · GPT Live Transcribe · ElevenLabs · Google · Mistral

MW

Modelwire Editorial

This synthesis and analysis was prepared by the Modelwire editorial team. We use advanced language models to read, ground, and connect the day’s most significant AI developments, providing original strategic context that helps practitioners and leaders stay ahead of the frontier.

Modelwire summarizes, we don’t republish. The Decoder originally reported this story as GPT Transcribe improves on its predecessor but can't catch ElevenLabs, Google, or Mistral on error rates”. The full content lives on the-decoder.com. If you’re a publisher and want a different summarization policy for your work, see our takedown page.

OpenAI launches speech models but lags rivals on accuracy · Modelwire