OPEA Audio-Speech-Recognition microservice for GenAI applications
10K+
ASR (Audio-Speech-Recognition) microservice helps users convert speech to text. When building a talking bot with LLM, users will need to convert their audio inputs (What they talk, or Input audio from other sources) to text, so the LLM is able to tokenize the text and generate an answer. This microservice is built for that conversion stage.
For detailed, step-by-step instructions on how to deploy the ASR microservice using Docker Compose on different Intel platforms, please refer to the deployment guide. The guide contains all necessary steps, including building images, configuring the environment, and running the service.
| Platform | Deployment Method | Link |
|---|---|---|
| Intel Xeon/Gaudi2 | Docker Compose | Deployment Guide |
| Intel Core | Docker Compose | Deployment Guide |
The following configurations have been validated for the ASR microservice.
| Deploy Method | Core Models | Platform |
|---|---|---|
| Docker Compose | Whisper | Intel Xeon/Gaudi2 |
| Docker Compose | Paraformer | Intel Core |
Content type
Image
Digest
sha256:9734fbc36…
Size
594 MB
Last updated
6 months ago
docker pull opea/asr