Video to Text
Transcribe a video's speech to text and subtitles (SRT/VTT) with Whisper — in your browser, nothing uploaded.
Pilih fail audio atau video
MP3, WAV, M4A, MP4, MOV, WebM — ditranskripsi dalam penyemak imbas anda dengan Whisper
Klik untuk memilih, atau lepaskan fail di sini
Cara Video to Text
- 1
Add your video
Drop in an MP4, MOV or WebM with an audio track — it's read in your browser, never uploaded.
- 2
Pick a model and language
Base is more accurate, Tiny is faster. Leave the language on Auto-detect or set it. The model downloads once (~40–80 MB) then it's cached.
- 3
Transcribe and export
Click Transcribe; when it's done, copy or download the transcript as plain text, SRT or VTT subtitles.
Related PDF tools
Video to Text — soalan lazim
Is my video uploaded?+
No. The transcription model (OpenAI's Whisper) runs entirely in your browser via WebAssembly. The video never leaves your device.
Why is there a one-time download?+
The Whisper model weights (~40 MB for Tiny, ~80 MB for Base, quantised) download the first time you use that size, then the browser caches them — after that it works offline and starts instantly.
How accurate is it?+
Base is solid for clear speech in a supported language; Tiny is noticeably rougher but much faster. Both struggle with heavy accents, crosstalk, music beds and poor audio — proofread the output.
What can I export?+
Plain text, or timed subtitles as SRT or VTT. The SRT/VTT timings come from the model's segment timestamps.
Adakah anda memuat naik fail saya ke pelayan?+
Tidak — setiap alat di sini memproses fail anda sepenuhnya pada peranti anda, dan tiada apa dimuat naik ke pelayan. Satu-satunya pengecualian ialah alat AI, yang menghantar teks yang diekstrak (bukan fail) ke Claude untuk melakukan kerjanya.
Adakah terdapat had saiz fail?+
Kami tidak mengenakan had buatan, tetapi kerana pemprosesan berlaku dalam pelayar anda, fail yang sangat besar (500+ MB atau 1,000+ muka surat) dihadkan oleh memori peranti anda, bukan oleh kami.
Advertisement
Advertisement