Run and serve large language models locally over an HTTP API (CPU inference).
6h
6.2K
The easiest way to get up and running with large language models.
1d
100M+
1.6K
Docker image for Ollama in Fred Hutch OCDO's WILDS, with OpenCode for LLM-driven coding workflows
4m
3.1K
Claude Code wired to Ollama Anthropic-compatible API, with configurable model.
3d
7.7K
https://github.com/dusty-nv/jetson-containers/packages/llm/ollama
1y
100K+
9
Fork of normal ollama with this fix: https://github.com/ollama/ollama/pull/6675
8m
2.4K
Ollama LLM server for NVIDIA Jetson (JetPack 6, arm64) with GPU acceleration.
3d
3.0K
For AccelBrain use. v0.1 builed from ollama v0.4.1. ; v0.2 builed from ollama v0.6.5.
10m
1.6K
An OpenAI API compatible server for local LLMs - llama2, mistral, codellama
3y
9.1K
14
Ollama image from this ollama Dockerfile : https://github.com/robertrosenbusch/gfx803_rocm
4m
967
ollama amd64 / CC 7.5 / T4 optimized / static q4_0 KV cache for 128k context in 16GB GPUs.
2y
2.3K
Ollama OCI/container image that contains Linux OS, Ollama, and Supervisor.
2m
981