Convert VTT to Plain Text
WebVTT wraps its words in a header, optional cue identifiers, NOTE blocks and positioning settings — none of which you want when what you are after is the transcript. This drops all of it and hands back the text. It runs in your browser, so a caption file you were sent in confidence stays on your machine.
Körs helt i din webbläsare – filen laddas aldrig upp.
Inget inläst ännu.
Behöver du undertexter direkt från videon? Ladda upp den och få en transkription på ~98 språk. 50 % rabatt första månaden med FIRST20.
Prova gratisVarför ScribeToAny
- Ladda upp MP3, MP4, M4A, MOV, AAC, WAV, OGG, OPUS, MPEG, WMA, WMV — ljudet extraheras automatiskt från video.
- Cirka 98 språk med automatisk identifiering, drivet av en motor i Whisper-klass på dedikerade GPU:er.
- Åtta exportformat: TXT, SRT, VTT, TSV, CSV, JSON, PDF och DOCX — plus ZIP-export i batch.
- Talarigenkänning märker varje segment. Klicka på valfri mening för att spela upp motsvarande ljud.
Vanliga frågor
What happens to NOTE and STYLE blocks?
They are skipped. The WebVTT spec forbids the cue arrow inside them, which is exactly how the parser tells them apart from real cues — nothing is guessed at.
Can I keep the timestamps?
Choose DOCX or PDF instead of TXT; both put a timestamp in front of each line. Plain text drops them by design, because the usual reason for wanting .txt is to paste the words somewhere else.
My captions came from YouTube and every line repeats. Why?
YouTube auto-captions roll each line into the next cue so the text scrolls on screen, so a straight export shows every line twice. Undoing that reliably needs the original audio — transcribing the video with ScribeToAny gives clean, non-overlapping segments instead.
Relaterade verktyg
Börja transkribera gratis
Gratisplan: 2 filer per dag, 100 minuter per månad. Inget kreditkort krävs.
Börja transkribera gratis