Intel® AI for Enterprise RAG Embedding Model Server powered by TorchServe
10K+
Part of the Intel® AI for Enterprise RAG (ERAG) ecosystem.
The OPEA ERAG TorchServe Embedding Model Server hosts embedding models using TorchServe, providing a scalable and efficient endpoint for generating vector embeddings from text or documents. It serves as the backend for the ERAG Embedding Microservice.
TorchServe is a lightweight, scalable, and easy-to-use model serving library for PyTorch models. It provides a RESTful API for serving trained models, allowing users to deploy and serve their models in production environments. Moreover, Torchserve supports Intel® Extension for PyTorch* for a performance boost on Intel-based Hardware.
This service integrates with other OPEA ERAG components:
OPEA ERAG is licensed under the Apache License, Version 2.0.
Copyright © 2024–2026 Intel Corporation. All rights reserved.
Content type
Image
Digest
sha256:34404acbc…
Size
2 GB
Last updated
5 months ago
docker pull opea/erag-torchserve_embedding