Transcribana
EN
EnglishFrançaisالعربية
Menu

Transcribana

Audio to text

Turn recordings into an editable transcript on your own device. Review your words and export text without uploading the audio.

Start transcribing
Start transcribing

Best supported on current desktop Chrome. Codec support, speed and memory vary by browser and device. Mobile transcription is not guaranteed.

Drop an audio or video file here

Interviews, meetings, lectures, and everything worth keeping.

Audio and video files are processed locally in your browser.

Supported file formats

MP3, WAV, M4A, AAC, FLAC, OGG, OPUS, MP4, MOV, WebM, MKV, M4V and 3GP with supported audio codecs. Files are checked locally. AVI, WMA and MPEG/MPG program streams are not supported. Codec support varies by browser.

Your files stay on your device.

Transcription runs locally in your browser. Your audio and video are not uploaded to our servers.

No account. No upload.
Ready when you are0:00 elapsed

Choose a recording to get started.

Progress will appear here.

Arabic uses a dedicated model with about 1.33 GB of weights and CPU/WASM processing. The first run may take several minutes. General model and GPU choices do not change the Arabic model.

Before you start

Start with the clearest recording you have. The app opens the primary audio track and checks whether your browser can decode it. Stereo is mixed to mono for recognition; speaker names are not assigned automatically.

A workflow that fits your file

Use this workflow for interviews, lectures and podcast recordings. PCM WAV, MP3, M4A/AAC, OGG and FLAC have tested decoding paths in desktop Chrome. Keep the original recording for checking names and quotations.

Review and use the result

After transcription, edit the readable document or review local correction suggestions. Copy Text and TXT follow your selected text version. SRT, VTT and JSON preserve the original segments.

Which audio files can I transcribe?

MP3 and PCM WAV are useful starting points. M4A is a container: a typical AAC voice memo differs from an unsupported codec in the same extension. AAC, FLAC and OGG decoding depends on the actual codec and browser. Choose the original file first; converting formats cannot restore unclear speech. If opening fails, read the error before attempting another export.

For a phone note, follow the exported voice-memo workflow. For long teaching recordings, see lecture review and resource limits. MP3-to-TXT, WAV-to-text and M4A-to-text use this same local file workflow; they are not separate recognition engines.

How it works

  1. Choose a recording

    Open an audio or video file. We check its audio track in your browser.

  2. Transcribe on your device

    Choose the spoken language or Auto Detect. Allow the required model to load.

  3. Review, then export

    Listen back to important passages, edit your text and download the format you need.

Your recording never needs to leave your device.

Speech recognition and text checking run locally. The website and model files still need network access. Hosting and model providers can receive ordinary connection information; they do not receive your recording through this tool.

Understand local processing

Questions, answered

Does my recording get uploaded?

No. The browser reads and transcribes it locally. Website assets and model downloads use the network separately.

Can I use it offline?

A model already loaded in an open page can be reused offline. Reloading, switching models or clearing storage may require network access.

Will text corrections change my subtitles?

No. Copy Text and TXT use the selected document version. SRT, VTT and JSON keep the original timed transcription.

Related tools

Useful guides

Ready to put your words to work?

Start transcribing

Your privacy choices

No analytics or advertising services are installed. There are no optional tracking categories to enable. The site stores your explicit language preference; browsers may cache models and website files.