Turn any recording into an editable script
Upload audio or video. Scenecast detects speakers, music, sound effects, and scenes — then lets you edit text, merge or split segments, rename speakers, and export.
Drop an audio or video file
MP3, WAV, M4A, MP4, MOV, WEBM — up to 1 GB
Your file is streamed to the transcription engine and not stored.
What you get
Speaker diarization
Each voice is separated and color-coded so you can follow multi-person recordings.
Music & sound effects
Non-speech moments — laughter, applause, music, and more — are tagged inline.
Edit, merge & split
Fix wording, rename speakers, and merge or split segments at any word boundary.
Export anywhere
Downloads reflect your edits: clean TXT, subtitle-ready SRT, or structured JSON.