Convert VTT to Plain Text

WebVTT wraps its words in a header, optional cue identifiers, NOTE blocks and positioning settings — none of which you want when what you are after is the transcript. This drops all of it and hands back the text. It runs in your browser, so a caption file you were sent in confidence stays on your machine.

Kjører helt i nettleseren din – filen lastes aldri opp.

Ingenting lastet inn ennå.

Trenger du undertekster rett fra videoen? Last den opp og få en transkripsjon på ~98 språk. 50 % rabatt første måned med FIRST20.

Prøv gratis

Hvorfor ScribeToAny

  • Last opp MP3, MP4, M4A, MOV, AAC, WAV, OGG, OPUS, MPEG, WMA, WMV — lyden trekkes automatisk ut fra video.
  • Rundt 98 språk med automatisk gjenkjenning, drevet av en motor i Whisper-klasse på dedikerte GPU-er.
  • Åtte eksportformater: TXT, SRT, VTT, TSV, CSV, JSON, PDF og DOCX — pluss ZIP-eksport i batch.
  • Talergjenkjenning merker hvert segment. Klikk på en setning for å spille av tilhørende lyd.

Ofte stilte spørsmål

What happens to NOTE and STYLE blocks?

They are skipped. The WebVTT spec forbids the cue arrow inside them, which is exactly how the parser tells them apart from real cues — nothing is guessed at.

Can I keep the timestamps?

Choose DOCX or PDF instead of TXT; both put a timestamp in front of each line. Plain text drops them by design, because the usual reason for wanting .txt is to paste the words somewhere else.

My captions came from YouTube and every line repeats. Why?

YouTube auto-captions roll each line into the next cue so the text scrolls on screen, so a straight export shows every line twice. Undoing that reliably needs the original audio — transcribing the video with ScribeToAny gives clean, non-overlapping segments instead.

Relaterte verktøy

Begynn å transkribere gratis

Gratisplan: 2 filer per dag, 100 minutter per måned. Ingen kredittkort kreves.

Begynn å transkribere gratis