Kubernetes generative AI operator; part of the Prem Operator platform
427
Deploy AI models and apps to Kubernetes without hitting every pitfall.
Deploy the operator and chat with Hermes in 2 or 3 steps. Note that you need a NVIDIA GPU with 8GB of RAM.
Install the NVIDIA Operator
Install the Prem Operator
$ helm install latest oci://registry-1.docker.io/premai/prem-operator-chart
Deploy the Hermes + Big-AGI example and forward the ports
$ kubectl apply -f examples/big-agi.yaml
$ kubectl port-forward services/big-agi-service 3000:3000
If you browse to localhost:3000 you'll be able to begin chatting as soon as LocalAI downloads the model.
https://github.com/richiejp/prem-operator/assets/988098/0f06b254-a1a0-4ae5-815a-ed84998f5c89โ
We'd love to hear from you! Feel free to create an issue or reach out to us on Prem's Discordโ on the operator channel.
Operator SDK is under Apache 2.0 license. See the LICENSEโ file for details.
Content type
Image
Digest
sha256:8d70dee91โฆ
Size
23.8 MB
Last updated
over 2 years ago
docker pull premai/prem-operator