Convert VTT to Plain Text

WebVTT wraps its words in a header, optional cue identifiers, NOTE blocks and positioning settings — none of which you want when what you are after is the transcript. This drops all of it and hands back the text. It runs in your browser, so a caption file you were sent in confidence stays on your machine.

全部在你的瀏覽器裡完成,檔案不會上傳。

還沒有內容。

需要直接從影片產生字幕?上傳影片即可,支援約 98 種語言。 用優惠碼 FIRST20,首月 5 折。

免費試用

為什麼選 ScribeToAny

  • 支援上傳 MP3、MP4、M4A、MOV、AAC、WAV、OGG、OPUS、MPEG、WMA、WMV,影片自動提取音軌。
  • 約 98 種語言,自動檢測,基於專用 GPU 上的 Whisper 級引擎。
  • 8 種匯出格式:TXT、SRT、VTT、TSV、CSV、JSON、PDF、DOCX,還支援批次打包 ZIP。
  • 說話人識別為每段標註發言者,點選任意句子即可回放對應原聲。

常見問題

What happens to NOTE and STYLE blocks?

They are skipped. The WebVTT spec forbids the cue arrow inside them, which is exactly how the parser tells them apart from real cues — nothing is guessed at.

Can I keep the timestamps?

Choose DOCX or PDF instead of TXT; both put a timestamp in front of each line. Plain text drops them by design, because the usual reason for wanting .txt is to paste the words somewhere else.

My captions came from YouTube and every line repeats. Why?

YouTube auto-captions roll each line into the next cue so the text scrolls on screen, so a straight export shows every line twice. Undoing that reliably needs the original audio — transcribing the video with ScribeToAny gives clean, non-overlapping segments instead.

相關工具

免費開始轉錄

免費額度:每天 2 個檔案、每月 100 分鐘,無需綁卡。

免費開始轉錄