About Whisper Web
Whisper Web transcribes audio and video in 99 languages using Whisper AI. Free mode runs locally in the browser with WebGPU or WebAssembly, requires no account, and exports TXT, SRT, VTT, or JSON. Optional paid cloud transcription supports longer files and batch uploads.
Whisper Web is a browser-based speech-to-text tool for turning interviews, meetings, lectures, podcasts, and other recordings into text. It uses Whisper AI and supports 99 languages with automatic language detection.
The free plan processes audio locally in the browser using WebGPU or WebAssembly, so recordings stay on the device. No installation, API key, or account is required for local transcription. Free files can be up to 200 MB and 20 minutes long. Transcripts can be exported as TXT, SRT, VTT, or JSON, including subtitle formats.
An optional Unlimited plan provides cloud transcription for longer files and batch uploads. Cloud processing sends selected files to the service, unlike the free local mode. It supports files up to 10 hours and 5 GB, with batches of up to 50 files. Pricing and current limits are available on the Whisper Web website.
The free plan processes audio locally in the browser using WebGPU or WebAssembly, so recordings stay on the device. No installation, API key, or account is required for local transcription. Free files can be up to 200 MB and 20 minutes long. Transcripts can be exported as TXT, SRT, VTT, or JSON, including subtitle formats.
An optional Unlimited plan provides cloud transcription for longer files and batch uploads. Cloud processing sends selected files to the service, unlike the free local mode. It supports files up to 10 hours and 5 GB, with batches of up to 50 files. Pricing and current limits are available on the Whisper Web website.