Audio & video to text & summary — automatic speech transcription, key point extraction, and speech cleanup
Or select any audio / video file (OGG, MP3, WAV, FLAC, M4A, MP4, WebM)
Lecture, meeting, or voice recording is processed directly on your device.
The first transcription downloads the selected offline Whisper model once, then it is cached. Your file never leaves your device.
📖 Vocabulary of specific terms and names (optional)
Without it, rare names, titles and professional terms are recognized as similar everyday words: a colleague's name breaks into two common words, a brand turns into a random noun, a lecture term into familiar nonsense. List such words here, separated by commas, and the model will recognize them more accurately. The rest of the speech is unaffected: nothing is dropped or forced to match the list.
Quick set:
20–30 words work better than a long list: the model reads only the beginning of the prompt and drops the rest.
A name or term recognized incorrectly?
Is it free to use?
Yes, completely free and no sign-up needed. The tool runs entirely in your browser: the file never goes to a server, so the server spends neither CPU nor traffic — there is nothing to charge for. No watermarks, no limits on file size or count.