Upload a WAV, MP3, MP4, or MOV file to test the transcription engine
The example below uses the OpenAI Whisper model to transcribe a myriad of hospital video to be accessibility compliant. This model one runs solely on my own machine so there's no additional licensing or usage costs, but please be patient as the model can take a couple of minutes to process.
File and bandwidth limits: 100MB and 10 minutes
This is a Flask Python implementation and hosted with Nginx.
Select a media file to upload. Supported formats: MP3, WAV, MP4, MOV.
A 35-second audio clip from Conan The Barbarian with multiple speakers and heavy accents