Choose a recording
Add a supported audio or video file, or record directly from your microphone.
Add an MP4, WebM, or other supported video and turn its speech into editable text, SRT captions, or WebVTT.
Choose from the complete Whisper language list or let the tool detect the spoken language automatically. Support includes English, Spanish, Chinese, Arabic, Hindi, Japanese, Russian, Vietnamese, and many more.
This tool uses the open-source Whisper speech recognition model to transcribe audio and video locally. LocalScribe is not affiliated with OpenAI. Learn about Whisper transcription →
Your file stays private and is processed on your device.
Fast transcription is selected by default.
Add a supported audio or video file, or record directly from your microphone.
Select the spoken language or use automatic detection, then follow the progress on screen.
Correct the text and save it as TXT, SRT, or WebVTT subtitles.
Extract the words from interviews, presentations, lessons, and social videos without sending the original file to a transcription server.
Capture spoken answers as editable text for research, reporting, and quotations.
Create a transcript that can become captions, descriptions, articles, and social posts.
Turn presentations, webinars, and lessons into searchable study material.
Download timestamped subtitle formats after reviewing the generated transcript.
Your selected recording and completed transcript stay on your device. The transcription tool does not upload them to our servers.
Choose a supported video file above, select the spoken language, and start transcription. Review the result and download TXT, SRT, or VTT.
Not currently. This tool accepts local files only. For content you own or have permission to use, save an authorized copy and select the file.
Yes. Select Arabic from the language menu or try automatic language detection.
MP4 and WebM commonly work. Other formats depend on the media decoders available in your browser and operating system.