ray-llm is an LLM serving solution to deploy and manage open source LLMs using Ray.
50K+
This is the publicly available set of Docker images for Anyscale/Ray's RayLLM (formerly Aviary) project.
RayLLM is an LLM serving solution that makes it easy to deploy and manage a variety of open source LLMs. It does this by:
| Name | Notes |
|---|---|
:0.3.1 | Release v0.3.1 |
:0.3.0 | Release v0.3.0 |
:latest | Most recently pushed version release image |
See: ray-project/ray-llm "Deploying RayLLM" for full instructions
Requires a machine with compatible NVIDIA A10 GPU and valid HUGGING_FACE_HUB_TOKEN to run the Amazon LightGPT model:
docker run \
--gpus all \
-e HUGGING_FACE_HUB_TOKEN=<your_token> \
--shm-size 1g \
-p 8000:8000 \
--entrypoint aviary \
anyscale/aviary:latest run --model models/continuous_batching/amazon--LightGPT.yaml
Source is available at https://github.com/ray-project/ray-llm
Content type
Image
Digest
sha256:ce052580e…
Size
11 GB
Last updated
about 9 hours ago
docker pull anyscale/ray-llm:nightly-py312-cu130