Convert VTT to Plain Text
WebVTT wraps its words in a header, optional cue identifiers, NOTE blocks and positioning settings — none of which you want when what you are after is the transcript. This drops all of it and hands back the text. It runs in your browser, so a caption file you were sent in confidence stays on your machine.
모두 브라우저에서 처리됩니다. 파일은 업로드되지 않습니다.
아직 불러온 내용이 없습니다.
영상에서 바로 자막을 만들고 싶나요? 업로드하면 약 98개 언어로 전사해 드립니다. 코드 FIRST20로 첫 달 50% 할인.
무료로 사용해 보기왜 ScribeToAny인가
- MP3, MP4, M4A, MOV, AAC, WAV, OGG, OPUS, MPEG, WMA, WMV 업로드 가능 — 비디오에서 오디오가 자동 추출됩니다.
- 전용 GPU의 Whisper급 엔진으로 자동 감지와 함께 약 98개 언어 지원.
- 8가지 내보내기 형식: TXT, SRT, VTT, TSV, CSV, JSON, PDF, DOCX — 일괄 ZIP 내보내기도 지원.
- 화자 인식이 모든 세그먼트에 레이블을 붙입니다. 문장을 클릭하면 해당 오디오를 다시 들을 수 있습니다.
자주 묻는 질문
What happens to NOTE and STYLE blocks?
They are skipped. The WebVTT spec forbids the cue arrow inside them, which is exactly how the parser tells them apart from real cues — nothing is guessed at.
Can I keep the timestamps?
Choose DOCX or PDF instead of TXT; both put a timestamp in front of each line. Plain text drops them by design, because the usual reason for wanting .txt is to paste the words somewhere else.
My captions came from YouTube and every line repeats. Why?
YouTube auto-captions roll each line into the next cue so the text scrolls on screen, so a straight export shows every line twice. Undoing that reliably needs the original audio — transcribing the video with ScribeToAny gives clean, non-overlapping segments instead.
관련 도구
무료로 전사 시작하기
무료 플랜: 하루 2개 파일, 월 100분. 신용카드 불필요.
무료로 전사 시작하기