Convert SMI to SRT

SAMI is Microsoft’s old caption format, and still what a lot of archived Korean broadcast files ship with. It records only the moment a caption should appear and leaves the ending implied by whatever comes next, which is why naive converters produce subtitles that never disappear. This works the endings out from the following sync point, drops the HTML wrapper, and writes a clean SubRip file.

すべてブラウザ内で処理されます。ファイルはアップロードされません。

まだ何も読み込まれていません。

動画から直接字幕を作りたいですか?アップロードすれば約 98 言語で文字起こしできます。 クーポン FIRST20 で初月 50% オフ。

無料で試す

ScribeToAny が選ばれる理由

  • MP3、MP4、M4A、MOV、AAC、WAV、OGG、OPUS、MPEG、WMA、WMV をアップロード可能 — 動画からは音声を自動抽出します。
  • 専用 GPU 上の Whisper クラスのエンジンにより、自動検出付きで約 98 言語に対応。
  • 8 つのエクスポート形式: TXT、SRT、VTT、TSV、CSV、JSON、PDF、DOCX — さらに一括 ZIP エクスポートにも対応。
  • 話者認識がすべてのセグメントにラベルを付けます。任意の文をクリックすると対応する音声を再生できます。

よくある質問

Why do SMI files not have end times?

Because the format marks moments, not spans: each sync point says "show this now", and the caption stays until the next one. Subtitling tools express a gap by emitting a sync whose paragraph is empty, so this converter treats a blank block as the previous caption’s end rather than as a subtitle of its own.

My SMI file has two languages in it — what happens?

SAMI can carry several language tracks in one file, tagged by a Class attribute on each paragraph. This converter reads every sync block in order, so a dual-language file gives you both languages inside the same cue. If you need one of them on its own, strip the other Class before converting.

The Korean text comes out as garbled characters.

Older SMI files are often saved in EUC-KR rather than UTF-8, and the browser reads dropped files as UTF-8. Re-save the .smi as UTF-8 in a text editor first — most editors offer this under "Save with encoding" — and the characters will come through correctly.

関連ツール

無料で文字起こしを始める

無料プラン: 1 日 2 ファイル、月 100 分。クレジットカード不要。

無料で文字起こしを始める