Convert VTT to Plain Text
WebVTT wraps its words in a header, optional cue identifiers, NOTE blocks and positioning settings — none of which you want when what you are after is the transcript. This drops all of it and hands back the text. It runs in your browser, so a caption file you were sent in confidence stays on your machine.
すべてブラウザ内で処理されます。ファイルはアップロードされません。
まだ何も読み込まれていません。
動画から直接字幕を作りたいですか?アップロードすれば約 98 言語で文字起こしできます。 クーポン FIRST20 で初月 50% オフ。
無料で試すScribeToAny が選ばれる理由
- MP3、MP4、M4A、MOV、AAC、WAV、OGG、OPUS、MPEG、WMA、WMV をアップロード可能 — 動画からは音声を自動抽出します。
- 専用 GPU 上の Whisper クラスのエンジンにより、自動検出付きで約 98 言語に対応。
- 8 つのエクスポート形式: TXT、SRT、VTT、TSV、CSV、JSON、PDF、DOCX — さらに一括 ZIP エクスポートにも対応。
- 話者認識がすべてのセグメントにラベルを付けます。任意の文をクリックすると対応する音声を再生できます。
よくある質問
What happens to NOTE and STYLE blocks?
They are skipped. The WebVTT spec forbids the cue arrow inside them, which is exactly how the parser tells them apart from real cues — nothing is guessed at.
Can I keep the timestamps?
Choose DOCX or PDF instead of TXT; both put a timestamp in front of each line. Plain text drops them by design, because the usual reason for wanting .txt is to paste the words somewhere else.
My captions came from YouTube and every line repeats. Why?
YouTube auto-captions roll each line into the next cue so the text scrolls on screen, so a straight export shows every line twice. Undoing that reliably needs the original audio — transcribing the video with ScribeToAny gives clean, non-overlapping segments instead.
関連ツール
無料で文字起こしを始める
無料プラン: 1 日 2 ファイル、月 100 分。クレジットカード不要。
無料で文字起こしを始める