To run locally:
docker run --gpus=all --shm-size 64g -p 7860:7860 -v ${HOME}/.cache:/root/.cache --rm 6zar/alpaca generate.py --load_8bit --base_model 'decapoda-research/llama-7b-hf' --lora_weights 'tloen/alpaca-lora-7b'
If you want to use with GPU, from my tests you need at least 20GB of VRAM at the GPU.
For this test you can use runpod.io to allocate container with GPU of your choice for a good price.
Create a template and set:
6zar/alpacagenerate.py --load_8bit --base_model 'decapoda-research/llama-7b-hf' --lora_weights 'tloen/alpaca-lora-7b'30GB7860,Content type
Image
Digest
sha256:a22a344e1…
Size
12.5 GB
Last updated
about 3 years ago
docker pull 6zar/alpaca