Convert VTT to Plain Text
WebVTT wraps its words in a header, optional cue identifiers, NOTE blocks and positioning settings — none of which you want when what you are after is the transcript. This drops all of it and hands back the text. It runs in your browser, so a caption file you were sent in confidence stays on your machine.
全部在你的浏览器里完成,文件不会上传。
还没有内容。
需要直接从视频生成字幕?上传视频即可,支持约 98 种语言。 用优惠码 FIRST20,首月 5 折。
免费试用为什么选 ScribeToAny
- 支持上传 MP3、MP4、M4A、MOV、AAC、WAV、OGG、OPUS、MPEG、WMA、WMV,视频自动提取音轨。
- 约 98 种语言,自动检测,基于专用 GPU 上的 Whisper 级引擎。
- 8 种导出格式:TXT、SRT、VTT、TSV、CSV、JSON、PDF、DOCX,还支持批量打包 ZIP。
- 说话人识别为每段标注发言者,点击任意句子即可回放对应原声。
常见问题
What happens to NOTE and STYLE blocks?
They are skipped. The WebVTT spec forbids the cue arrow inside them, which is exactly how the parser tells them apart from real cues — nothing is guessed at.
Can I keep the timestamps?
Choose DOCX or PDF instead of TXT; both put a timestamp in front of each line. Plain text drops them by design, because the usual reason for wanting .txt is to paste the words somewhere else.
My captions came from YouTube and every line repeats. Why?
YouTube auto-captions roll each line into the next cue so the text scrolls on screen, so a straight export shows every line twice. Undoing that reliably needs the original audio — transcribing the video with ScribeToAny gives clean, non-overlapping segments instead.
相关工具
免费开始转录
免费额度:每天 2 个文件、每月 100 分钟,无需绑卡。
免费开始转录