Transcribana
MP3 to text
Transcribe an MP3 recording into text, with local processing and downloadable subtitle files.
Start transcribingStart transcribing
Best supported on current desktop Chrome. Codec support, speed and memory vary by browser and device. Mobile transcription is not guaranteed.
Drop an audio or video file here
Interviews, meetings, lectures, and everything worth keeping.
Audio and video files are processed locally in your browser.
Supported file formats
MP3, WAV, M4A, AAC, FLAC, OGG, OPUS, MP4, MOV, WebM, MKV, M4V and 3GP with supported audio codecs. Files are checked locally. AVI, WMA and MPEG/MPG program streams are not supported. Codec support varies by browser.
Transcription runs locally in your browser. Your audio and video are not uploaded to our servers.
Advanced settings
Models download on first use. A loaded model is reused; change model or processing mode to release it.
Choose a recording to get started.
Progress will appear here.Your words, ready to use.
Document edits apply to copied text and TXT. Subtitles and JSON preserve the original timed transcription.
Processing details
Timestamps are model estimates. Total RAM and GPU memory are not available to this page; use your browser’s task manager.
Arabic uses a dedicated model with about 1.33 GB of weights and CPU/WASM processing. The first run may take several minutes. General model and GPU choices do not change the Arabic model.
Before you start
Select the original MP3 rather than re-encoding it first. Converting a low-quality MP3 to WAV does not restore lost speech detail. The app checks and decodes your file locally.
A workflow that fits your file
For podcast excerpts and voice recordings, choose the spoken language explicitly when you know it. Background music, heavy compression and overlapping voices can affect recognition. There is no automatic speaker labeling.
Review and use the result
Use TXT for notes and quotes after checking the recording. Keep JSON if you need segment times in another workflow. Processing speed depends on your hardware, the model and the length of the recording.
From an MP3 recording to a TXT document
Choose your MP3, select the spoken language and start transcription. Review names and numbers, make any edits in the readable transcript, then select Download TXT. This exports plain text, not a converted audio file. You can also copy the selected text version.
A low-bitrate MP3 may already have lost speech detail; exporting it as WAV will not recover that detail. For a finished podcast episode, check music transitions and guest names. If the source is a voice memo in M4A instead, use the audio-format guidance rather than converting solely to obtain an MP3 extension.
How it works
Choose a recording
Open an audio or video file. We check its audio track in your browser.
Transcribe on your device
Choose the spoken language or Auto Detect. Allow the required model to load.
Review, then export
Listen back to important passages, edit your text and download the format you need.
Your recording never needs to leave your device.
Speech recognition and text checking run locally. The website and model files still need network access. Hosting and model providers can receive ordinary connection information; they do not receive your recording through this tool.
Understand local processingQuestions, answered
Does my recording get uploaded?
No. The browser reads and transcribes it locally. Website assets and model downloads use the network separately.
Can I use it offline?
A model already loaded in an open page can be reused offline. Reloading, switching models or clearing storage may require network access.
Will text corrections change my subtitles?
No. Copy Text and TXT use the selected document version. SRT, VTT and JSON keep the original timed transcription.
Related tools
Useful guides
How to turn an audio or video file into a usable transcript
A step-by-step guide to local transcription: check the audio track, choose a language, review the words and export text or subtitles.
Transcription accuracy: prepare the source and check the words
Prepare an audio or video file, choose a language, review recognition errors and export the right version of your transcript.