Skip to main content
All posts
Guides

How to Transcribe Audio with AI

A practical tutorial for turning recordings into text with MAI Transcribe 2 and linking back to the online speech to text tool.

Transcribing audio with AI is now a three-step job: get a clean-enough file, pick a speech to text model, and export text you can edit.

1. Prepare the file

Export MP3, WAV, M4A, or FLAC. Short clips beat hour-long dumps when you are testing a new model. Remove obvious silence if your editor makes that easy.

2. Run MAI Transcribe 2

Open the homepage tool and upload the file. MAI Transcribe 2 detects language automatically and can label speakers. You do not need an Azure login for that first try.

3. Clean and publish

Copy the transcript, fix names, and add headings. If you need captions, keep the timestamps. If you need notes, switch to a clean reading copy.

For a model choice, read MAI Transcribe 2 vs Whisper. For scenes like meetings and podcasts, use the speech to text page.