Convert Opus to Text
Opus is the codec behind most voice notes and calls — WhatsApp, Telegram, Discord — which means a lot of what people actually say to each other is locked in .opus files you cannot search. Upload the voice note and ScribeToAny transcribes it into timestamped text you can read, quote and export, with optional speaker labels for a two-person exchange. Works on every plan.
How it works
- 1
Upload
Drop in an audio or video file, or paste a link. All common formats are accepted.
- 2
Transcribe
AI transcribes with timestamps and optional speaker labels, usually in minutes.
- 3
Export
Edit segments online, then export TXT, SRT, VTT, TSV, CSV, JSON, PDF or DOCX.
Why ScribeToAny
- Upload MP3, MP4, M4A, MOV, AAC, WAV, OGG, OPUS, MPEG, WMA, WMV — audio is extracted from video automatically.
- About 98 languages with automatic detection, powered by a Whisper-class engine on dedicated GPUs.
- Eight export formats: TXT, SRT, VTT, TSV, CSV, JSON, PDF and DOCX — plus batch ZIP export.
- Speaker recognition labels every segment. Click any sentence to replay the matching audio.
Frequently asked questions
Can I transcribe WhatsApp or Telegram voice notes?
Yes — export or save the voice note as an .opus (or .ogg) file and upload it here. The audio is transcribed as-is; there is no need to convert it first, and no app to install.
Does it separate two speakers?
Turn on speaker detection and a two-person voice note comes back as a labelled back-and-forth rather than one block. It is a best-effort split by voice; rename the speakers in the editor.
What can I do with the transcript?
Edit it in the browser, click a line to replay that moment, then export TXT, SRT, VTT, TSV, CSV, JSON, PDF or DOCX. On the Max plan you can also translate it into any of 56 languages.
Related tools
Start transcribing free
Free plan: 2 files per day, 100 minutes per month. No credit card required.
Start transcribing free