Muse Voice
Muse Voice is an online speech-to-text workspace that transcribes audio and video with speaker labels and timestamps, then exports TXT, DOCX, PDF, SRT or VTT.
About Muse Voice
Muse Voice is an online speech-to-text workspace that takes a recording from intake to a reviewed handoff in the browser. Creators, journalists, students, and teams use Muse Voice to transcribe meetings, interviews, podcasts, lectures, and video files, then correct the draft before they publish captions or share notes.
Upload MP3, WAV, M4A, FLAC, MP4, or MOV files up to 1 GB, record live in the page, or paste a hosted media URL. Muse Voice speech to text can auto-detect or use a chosen language across about 100 languages. Treat every automatic result as a draft.
Key Features
- File upload, in-browser recording, or a hosted media URL
- Optional speaker separation with rename-once labels
- Word-level timestamps with click-to-replay
- An editor for search, name fixes, and speaker cleanup
- Six exports: TXT, DOCX, PDF, SRT, VTT, and JSON
Use Cases
Meetings become decision logs, interviews keep named quotations, podcasts become show notes, lectures become study notes, and videos become SRT or VTT captions.
Pricing
New accounts can verify and receive 5 transcription minutes. Paid plans add monthly or yearly minute allowances. Details: https://musevoice.pro/pricing.