The news, 365 days behind — on purpose Delayed live · replaying 2025

One Year Ago.AI

Remember how fast this is.

23SEPT2022replayed
one year on
researchOpenAI

OpenAI open-sources Whisper, a multilingual speech recognition model claiming human-level robustness

The model, trained on 680,000 hours of weakly supervised internet audio, is released for researchers and developers, with early adopters already building local transcription tools for journalists.

OpenAI today releases Whisper, an open-source neural network for automatic speech recognition, along with models and inference code. The model was trained on 680,000 hours of multilingual and multitask audio data collected from the internet, using weak supervision — meaning it simply learned to predict transcripts without explicit alignment or language labels. According to the accompanying paper, Whisper approaches human-level robustness and accuracy on English speech recognition and can automatically transcribe and translate languages including Spanish, Italian, and Japanese, all in a zero-shot setting without fine-tuning.

The Verge reported that the model installs easily on a local machine and can transcribe audio privately without sending files to the cloud. Journalist Peter Sterne and developer Christina Warren announced they are building Stage Whisper, a free transcription app based on Whisper, aiming to give journalists a secure alternative to cloud services. Sterne told The Verge the model produced the best transcription he had used aside from human transcribers.

The release marks a departure from OpenAI’s recent cautious approach to model access: the company has limited access to GPT-3 and DALL-E, citing a desire to learn more about real-world use and continue to iterate on its safety systems. Whisper’s release may signal a different strategy for speech technology, but the blog post framing remains focused on research.

M
Mitchell Clark

The Verge reporter described installing Whisper as easy as running a single Terminal command and said it turned out to be even easier than I'd imagined, noting that a journalist and developer are teaming up to create a free, secure transcription app called Stage Whisper for journalists.

P
Peter Sterne

Sterne said of Whisper that it was the best transcription he had ever used, with the exception of human transcribers, and that he decided to create Stage Whisper because journalists need good auto-transcription apps today.

One year later — open only if you can handle spoilers

Whisper became the de facto open-source transcription backbone for a generation of startups and tools, from podcasting to medical dictation. Its MIT license and robust performance made it ubiquitous, while its 2022 release stands as OpenAI's last major open-weights release before shifting toward gated models and paid APIs.

Replay thisPost on XRedditHNLinkedIn

The Weekly Replay · free by email

This week, one year ago — every Sunday.

One email each Sunday: the week's replayed AI news, with the one-year-later annotations included. Written like it's breaking — dated like it isn't.

Free · double opt-in · unsubscribe anytime · privacy