Upload or record
Bring an authorized file or capture the conversation while it happens.
AUDIO TRANSCRIPTION FOR PUBLISHING WORK
Upload a podcast, interview, meeting or voice recording. CastTranscript keeps the transcript, speakers, subtitles, chapters and show notes in one private editor.
MP3, M4A, WAV, MP4, MOV or WebM
or drag and drop it anywhere hereOne working document
Correct a name, replay a sentence or relabel a speaker in the document. Playback, subtitle cues and generated notes stay attached to the same edit.
NoraWe kept coming back to the same question.
MayaIt is usually the detail you almost edit out. That is the line the listener remembers.
NoraLet’s move it into the opening chapter and keep the pause before it.
From source to deliverable
Bring an authorized file or capture the conversation while it happens.
Batch, live and fallback models stay separate, visible processing passes.
Play from a cue, correct the words and rename speakers in one document.
Download TXT, Markdown, SRT or VTT without publishing a public page.
Workflows by source
Each tool changes the input, review surface and expected output. They are not duplicate upload pages with different headlines.
Get the transcript, chapters and show notes needed for the episode page.
Podcast transcription 02Keep speaker turns, timestamps and quotes attached to the source recording.
Interview transcription 03Turn a finished meeting recording into searchable text and action-ready notes.
Meeting transcription 04Review sentence timing and export SRT or VTT when the source provides cues.
MP3 to SRT 05Upload authorized media or import caption text you can already access.
YouTube transcript 06Turn an authorized short video into editable text or captions.
TikTok transcriptCapability-aware routing
CastTranscript chooses a model for the job, records the model used and shows what came back. You keep one editor and one export workflow even as providers improve.
Compare current live modelsAudio transcription for real work
Convert MP3, M4A and WAV recordings into editable text, or turn an authorized video into a transcript and subtitle cues. The same private workspace handles podcasts, interviews, meetings and short social clips without forcing every job into one model or one output.
For long media, the service processes queued sections and merges verified timing back into a single document. When a provider returns sentence timing but not word timing or speaker labels, CastTranscript shows the limitation instead of manufacturing alignment with an unrelated model.
When the text is ready, copy it or export TXT, Markdown, SRT or WebVTT. Projects remain private to the account, and exports stay files rather than accidental public transcript pages.
Questions, answered
Upload common audio and video formats, record a live conversation or import caption text you are authorized to use. Each source opens as a private, editable project.
The current default is MAI-Transcribe-2 for uploaded media and Grok Voice Transcribe 2.0 for live sessions. Muse Voice Transcribe remains a configured live fallback. CastTranscript records the model that actually ran.
Yes when the project contains verified cue timing. Untimed caption imports remain available as TXT or Markdown rather than receiving invented timestamps.
Speaker labels appear when the selected model returns diarization data. You can rename them in the editor. Missing speaker data is disclosed without blocking the transcript.
No. Projects remain inside the owner’s account. CastTranscript supports copy and file export, not public transcript URLs.
Private by default
Start with 30 free minutes. Upload a file, record live, correct the transcript and export the result.