Community User
Displaying 1 to 15 of 15 repositories
llama.cpp for quantizing and inferencing large language models in C++
20d
10K+
Automate the Chromium web browser with Selenium and Python
11m
465
exllamav2 is "a fast inference library for running LLMs locally on modern consumer-class GPUs".
1y
1.3K
1