Transcribana
Generate subtitles from a recording
Create original-language subtitle files from audio or video, with estimated start and end times.
Start transcribingStart transcribing
Best supported on current desktop Chrome. Codec support, speed and memory vary by browser and device. Mobile transcription is not guaranteed.
Drop an audio or video file here
Interviews, meetings, lectures, and everything worth keeping.
Audio and video files are processed locally in your browser.
Supported file formats
MP3, WAV, M4A, AAC, FLAC, OGG, OPUS, MP4, MOV, WebM, MKV, M4V and 3GP with supported audio codecs. Files are checked locally. AVI, WMA and MPEG/MPG program streams are not supported. Codec support varies by browser.
Transcription runs locally in your browser. Your audio and video are not uploaded to our servers.
Advanced settings
Models download on first use. A loaded model is reused; change model or processing mode to release it.
Choose a recording to get started.
Progress will appear here.Your words, ready to use.
Document edits apply to copied text and TXT. Subtitles and JSON preserve the original timed transcription.
Processing details
Timestamps are model estimates. Total RAM and GPU memory are not available to this page; use your browser’s task manager.
Arabic uses a dedicated model with about 1.33 GB of weights and CPU/WASM processing. The first run may take several minutes. General model and GPU choices do not change the Arabic model.
Before you start
Transcribe the recording, then open the Subtitles view. It shows the original speech segments and model-estimated timestamps. Review cue boundaries against your video before publishing.
A workflow that fits your file
Export SRT for common editor workflows or VTT for web video. No extra inference is needed to switch between these formats. Segments with missing or invalid times are excluded from timed exports and remain in text and JSON.
Review and use the result
This is a starting subtitle track, not a finished caption-editing suite. Speaker labels, sound-effect descriptions, line-length styling and burned-in captions are not generated. Correct readable text separately; subtitle exports retain the original wording.
Create captions from speech, then review them in context
Select the final audio or video, transcribe, and open Subtitles to inspect recognized text with its times. Download SRT or VTT for the destination player. This produces same-language speech captions; it does not translate, identify speakers, label sound effects or burn text into the video.
Check whether a cue appears when its words are audible and stays readable. Recognition segments are not a professional caption layout. Document edits do not flow into these cues. The subtitle format guide explains what each export contains; podcast publishing guidance covers keeping the timing aligned to a final edit.
How it works
Choose a recording
Open an audio or video file. We check its audio track in your browser.
Transcribe on your device
Choose the spoken language or Auto Detect. Allow the required model to load.
Review, then export
Listen back to important passages, edit your text and download the format you need.
Your recording never needs to leave your device.
Speech recognition and text checking run locally. The website and model files still need network access. Hosting and model providers can receive ordinary connection information; they do not receive your recording through this tool.
Understand local processingQuestions, answered
Does my recording get uploaded?
No. The browser reads and transcribes it locally. Website assets and model downloads use the network separately.
Can I use it offline?
A model already loaded in an open page can be reused offline. Reloading, switching models or clearing storage may require network access.
Will text corrections change my subtitles?
No. Copy Text and TXT use the selected document version. SRT, VTT and JSON keep the original timed transcription.
Related tools
Useful guides
Arabic audio transcription: a practical review workflow
Choose the Arabic model, prepare a clear recording and review dialects, names and subtitle timing without uploading your media.
SRT or VTT: which subtitle file do you need?
Understand the practical difference between SRT and WebVTT, choose an export and check captions before publishing.