Choose a recording
Add a supported audio or video file, or record directly from your microphone.
Use Whisper AI to convert audio and video into editable text across 99 supported languages, directly from this web page.
Choose from the complete Whisper language list or let the tool detect the spoken language automatically. Support includes English, Spanish, Chinese, Arabic, Hindi, Japanese, Russian, Vietnamese, and many more.
This tool uses the open-source Whisper speech recognition model to transcribe audio and video locally. LocalScribe is not affiliated with OpenAI.
Your file stays private and is processed on your device.
Fast transcription is selected by default.
Add a supported audio or video file, or record directly from your microphone.
Select the spoken language or use automatic detection, then follow the progress on screen.
Correct the text and save it as TXT, SRT, or WebVTT subtitles.
Whisper is an open-source speech recognition model created by OpenAI. LocalScribe makes it available as a private online transcription tool and is not affiliated with OpenAI.
Transcribe English and multilingual recordings with automatic or manual language selection.
Use common recording and media formats that your web browser can decode.
Use online Whisper transcription without uploading your selected recording to us.
Correct the output and save it as TXT, SRT, or WebVTT.
Your selected recording and completed transcript stay on your device. The transcription tool does not upload them to our servers.
No. LocalScribe uses the open-source Whisper model but is not affiliated with, endorsed by, or operated by OpenAI.
Yes. There is no sign-up or per-minute fee. Your device performs the transcription work.
This version focuses on transcription in the spoken language. A separate translation workflow is not currently offered.
Fast is selected by default for quicker transcription and lower memory use. Choose Higher accuracy for important recordings when you can allow more processing time.