Add audio
Drop an existing file or capture a fresh microphone recording in the browser.
Drag in an audio file or record directly from your microphone. Your audio stays on your device while a browser-based Whisper model creates the transcript.
The first transcription downloads the selected AI model to the browser cache. Larger models can improve accuracy but need more memory and download time.
Select or record audio, choose your settings, then start transcription.
Loading the speech-recognition model.
Drop an existing file or capture a fresh microphone recording in the browser.
The selected Whisper model runs on your device using browser-based machine learning.
Review the transcript and save it as plain text, subtitles or structured JSON.
Your audio is processed in browser memory and is not intentionally uploaded by this website. Model files are downloaded from Hugging Face/CDN infrastructure and cached by your browser. Microphone recording requires permission and usually HTTPS. Transcription speed and accuracy depend on audio quality, language, model size and your device.