Skip to content

Overview

Whisper is a general-purpose speech recognition model. It is trained on a large dataset of diverse audio and is also a multitask model that can perform multilingual speech recognition as well as speech translation and language identification.

Join our Discord Community!

🎉 Connect with other users, get help, and stay updated on the latest features!
Join our Discord Server

Features¶

Current release (v1.9.1) supports following whisper models:

Quick Usage¶

docker run -d -p 9000:9000 -e ASR_MODEL=base -e ASR_ENGINE=openai_whisper onerahmet/openai-whisper-asr-webservice:latest
docker run -d --gpus all -p 9000:9000 -e ASR_MODEL=base -e ASR_ENGINE=openai_whisper onerahmet/openai-whisper-asr-webservice:latest-gpu

for more information:

Credits¶

  • This software uses libraries from the FFmpeg project under the LGPLv2.1