Convert VTT to Plain Text

WebVTT wraps its words in a header, optional cue identifiers, NOTE blocks and positioning settings — none of which you want when what you are after is the transcript. This drops all of it and hands back the text. It runs in your browser, so a caption file you were sent in confidence stays on your machine.

Chạy hoàn toàn trong trình duyệt của bạn — tệp không bao giờ được tải lên.

Chưa có nội dung nào.

Cần tạo phụ đề trực tiếp từ video? Tải lên và nhận bản ghi ở ~98 ngôn ngữ. Giảm 50% tháng đầu với mã FIRST20.

Dùng thử miễn phí

Tại sao chọn ScribeToAny

  • Tải lên MP3, MP4, M4A, MOV, AAC, WAV, OGG, OPUS, MPEG, WMA, WMV — âm thanh được trích xuất tự động từ video.
  • Khoảng 98 ngôn ngữ với tự động nhận diện, được hỗ trợ bởi động cơ cấp Whisper trên GPU chuyên dụng.
  • Tám định dạng xuất: TXT, SRT, VTT, TSV, CSV, JSON, PDF và DOCX — cùng xuất ZIP hàng loạt.
  • Nhận dạng người nói gắn nhãn mọi phân đoạn. Nhấp vào bất kỳ câu nào để phát lại đoạn âm thanh tương ứng.

Câu hỏi thường gặp

What happens to NOTE and STYLE blocks?

They are skipped. The WebVTT spec forbids the cue arrow inside them, which is exactly how the parser tells them apart from real cues — nothing is guessed at.

Can I keep the timestamps?

Choose DOCX or PDF instead of TXT; both put a timestamp in front of each line. Plain text drops them by design, because the usual reason for wanting .txt is to paste the words somewhere else.

My captions came from YouTube and every line repeats. Why?

YouTube auto-captions roll each line into the next cue so the text scrolls on screen, so a straight export shows every line twice. Undoing that reliably needs the original audio — transcribing the video with ScribeToAny gives clean, non-overlapping segments instead.

Công cụ liên quan

Bắt đầu phiên âm miễn phí

Gói miễn phí: 2 tệp mỗi ngày, 100 phút mỗi tháng. Không cần thẻ tín dụng.

Bắt đầu phiên âm miễn phí