Video to Text
Transcribe a video's speech to text and subtitles (SRT/VTT) with Whisper — in your browser, nothing uploaded.
Seleccioneu un fitxer d'àudio o vídeo
MP3, WAV, M4A, MP4, MOV, WebM — transcrit al vostre navegador amb Whisper
Fes clic per triar o deixa anar un fitxer aquí
Com Video to Text
- 1
Add your video
Drop in an MP4, MOV or WebM with an audio track — it's read in your browser, never uploaded.
- 2
Pick a model and language
Base is more accurate, Tiny is faster. Leave the language on Auto-detect or set it. The model downloads once (~40–80 MB) then it's cached.
- 3
Transcribe and export
Click Transcribe; when it's done, copy or download the transcript as plain text, SRT or VTT subtitles.
Related PDF tools
Video to Text — preguntes freqüents
Is my video uploaded?+
No. The transcription model (OpenAI's Whisper) runs entirely in your browser via WebAssembly. The video never leaves your device.
Why is there a one-time download?+
The Whisper model weights (~40 MB for Tiny, ~80 MB for Base, quantised) download the first time you use that size, then the browser caches them — after that it works offline and starts instantly.
How accurate is it?+
Base is solid for clear speech in a supported language; Tiny is noticeably rougher but much faster. Both struggle with heavy accents, crosstalk, music beds and poor audio — proofread the output.
What can I export?+
Plain text, or timed subtitles as SRT or VTT. The SRT/VTT timings come from the model's segment timestamps.
Pugeu els meus fitxers a un servidor?+
No: cada eina aquí processa els teus fitxers totalment al teu dispositiu, i no es puja res a un servidor. L'única excepció són les eines d'IA, que envien el text extret (no el fitxer) a Claude per fer la seva feina.
Hi ha un límit de mida de fitxer?+
No imposem cap límit artificial, però com que el processament passa al teu navegador, els fitxers molt grans (500+ MB o 1.000+ pàgines) estan limitats per la memòria del teu dispositiu, no per nosaltres.
Advertisement
Advertisement