Sign inSign up

opea/erag-vllm-audio

By opea

•Updated 3 months ago

Intel® AI for Enterprise RAG VLLM Audio Model Server for converting audio to text

Image
0

2.2K

opea/erag-vllm-audio repository overview

⁠VLLM Audio model server

Part of the Intel® AI for Enterprise RAG (ERAG) ecosystem.

⁠🔍 Overview

The VLLM Audio model server converts audio files to text using Whisper or compatible ASR models. It extends base vllm image to include audio related packages.

⁠Features
  • Converts audio files to text using Whisper ASR models
  • Supports multiple audio formats (WAV, MP3)
  • OpenAI-compatible API interface
  • Configurable language detection and model selection

This service integrates with other OPEA ERAG components:

  • OPEA ERAG ASR Microservice sends the requests to it to convert audio files to text.

⁠License

OPEA ERAG is licensed under the Apache License, Version 2.0.

Copyright © 2026 Intel Corporation. All rights reserved.

Tag summary

Content type

Image

Digest

sha256:3c1b5490e…

Size

1.9 GB

Last updated

5 months ago

docker pull opea/erag-vllm-audio