Sign inSign up

toolboc/nv-llama

By toolboc

•Updated over 3 years ago

NVIDIA Jetson Accelerated build of https://github.com/ggerganov/llama.cpp

Image
0

233

toolboc/nv-llama repository overview

⁠Download models into local folder then mount and run as shown:

docker run --rm -it --name llama --net=host --gpus all -v ~/src/llama.cpp/models:/models toolboc/nv-llama:r35.2.1 --run --model /models/13B/llama-13b.ggmlv3.q6_K.bin --n-predict 512 --n-gpu-layers 43 --repeat_penalty 1.0 --color --interactive-first

Tag summary

Content type

Image

Digest

sha256:d06e5799e…

Size

5.8 GB

Last updated

over 3 years ago

docker pull toolboc/nv-llama:r35.2.1