Your recordings stay with you
The recognition model runs on your device. Your selected audio and transcript are not sent to a cloud transcription service. Export what you want to keep.
Transcribe locally with Whisper, review your text, and export TXT, SRT or VTT for free. No signup. Your recordings stay on your device.
MP3 · WAV · M4A · MP4 · WEBM · MOV · OGG · OPUS · AAC · FLAC · Up to 200 MB / 20 minutes
A free browser workspace built around local speech recognition and practical subtitle tools.
Whisper is an open-source speech recognition model from OpenAI. It recognizes speech in multiple languages and turns it into text. This website brings file selection, transcript review, and subtitle export together around that model.
An interview, a lecture recording, or spoken dialogue in a video can become a transcript. Find a quotation, check a technical term, prepare notes, or use timed subtitles in your next edit.
Explore the official Whisper projectKeep your source files close while turning speech into something you can use.

The recognition model runs on your device. Your selected audio and transcript are not sent to a cloud transcription service. Export what you want to keep.
Open the page and choose a file. The first use downloads a public model; later tasks can reuse the browser cache while it remains available.
Local recognition uses memory and computing power. Prefer GPU acceleration when available, or use the tested CPU path. Larger files can take longer.
Full transcription has been tested in desktop Chrome on macOS. A responsive layout does not mean real-phone recognition has been verified. Start with a short, clear recording.
Recognize speech, review the result, and export the format you need.
Turn the words in a recording into original-language text. Check names, numbers, accents, and specialist terms.
Keep recordings in your browser with no account required. Web pages and model downloads still need a network connection.
Read the audio track of a compatible local video and organize words into timed segments. Preview captions without changing the original video.
Edit words and start/end seconds, replay a segment, then download TXT text or SRT/VTT subtitles.
The spoken language and the website language are separate choices.
An English podcast produces English text; a Spanish interview produces Spanish text. Choose the spoken language or automatic detection. Accuracy varies by language, accent, and model. You can also translate to English for free. Other target languages and dubbing are not supported.
Get a transcript, create captions, or convert an existing subtitle file.
Turn interviews, lectures, and voice notes into editable text with local Whisper transcription.
MP3 · WAV · M4A · WEBM · OGG · OPUS · AAC · FLAC → TXT / SRT / VTTTranscribe the spoken audio in a local video and export an editable text transcript.
MP4 · WEBM · MOV → TXT / SRT / VTTCreate original-language SRT or VTT subtitles from audio or video, then edit the words and timing.
MP4 · WEBM · MOV · MP3 · WAV · M4A · OGG · OPUS · AAC · FLAC → TXT / SRT / VTTTranscribe MP3 podcasts, interviews, and recordings into a downloadable text file or subtitles.
MP3 → TXT / SRT / VTTTranscribe WAV recordings from interviews, lessons, and recording equipment without uploading the audio.
WAV → TXT / SRT / VTTConvert an existing SRT subtitle file to WebVTT while preserving text and millisecond timing.
SRT → VTTConvert WebVTT subtitles into SRT for use in compatible editors and media players.
VTT → SRTExtract readable text from SRT subtitles without sequence numbers or timestamps.
SRT → TXTTurn recorded speech and voice notes into editable text for free. Record in the browser or select a local file, review the words, and export without an account.
MP3 · WAV · M4A · WebM · OGG · OPUS · AAC · FLAC → TXTConvert M4A voice memos into editable text or timed subtitles in your browser. Keep your recording local and export TXT, SRT, or VTT.
M4A → TXT / SRT / VTTTranscribe OGG voice recordings locally with Whisper. Review the transcript against the audio and download text or timed subtitles.
OGG → TXT / SRT / VTTExtract readable text from WebVTT subtitles for free. Remove timecodes and subtitle markup locally while preserving the words in order.
VTT → TXTConvert a saved YouTube video to text for free. Select a local MP4, WebM or MOV, review the speech, and export TXT, SRT or VTT. Direct video links are not supported.
MP4 · WebM · MOV → TXT / SRT / VTTSpend your attention on the content.
Start with a clear, short local audio or video file.
Compare the transcript with the source and correct important details.
Download plain text for notes or timed subtitles for a player or editor.
Practical workflows for podcast excerpts, meeting recordings, interviews, and lectures. Review locally, then export text or subtitles.
Turn a podcast recording into editable text and SRT or VTT captions. Transcribe locally, review guest names, and export material for your episode page.
Convert a saved meeting recording into searchable text without inviting a meeting bot. Review decisions and numbers locally, then export your transcript.
Transcribe interview recordings without uploading the audio. Replay passages, check quotations and speaker names, then export editable text or subtitles.
Turn recorded lectures into searchable text or original-language subtitles. Review technical terms against the audio and export notes you can study from.
Useful answers before you begin.
Yes. It runs multilingual Whisper tiny, base and small models locally through Transformers.js. Base is the default; tiny has a smaller download but needs careful review.
Local transcription, video audio recognition, basic subtitle editing, TXT/SRT/VTT export, and subtitle format conversion. No account is needed.
No. Selected files and results stay in the browser. The program and public models are downloaded. Download your result before leaving because there is no cloud project storage.
No. Transcription writes speech in its original language. Translation changes it into another language. Select Translate to English for English text and subtitles. Other target languages and voice replacement are not supported.
You can try, but noise, music, overlapping voices, and unusual names can cause errors. Use a clear original and check important content by listening.
Prepare your file, check the result, and download the format you need.
Prepare your device, understand the first model download, and keep a local transcript.
Choose the right source, set its language, proofread, and export text or captions.
Generate original-language captions, adjust timing, and take them into your player or editor.
Compare headers, timestamps, web styling, and what a format conversion keeps.
Prepare and review an SRT file locally, then upload it as a caption track for a video you manage in YouTube Studio.
Work within the 200 MB and 20-minute limits: prepare ordered excerpts, track time offsets, and preserve the original recording.
Understand the difference between website language, spoken-language recognition, and subtitle translation before starting Whisper.