(1)
Distilled LLaMA by DeepSeek, fast and optimized for real-world tasks
1y
100K+
79
mxbai-embed-large-v1 is a top English embed model by Mixedbread AI, great for RAG and more.
1y
10K+
3
Newest LLama 3 release with improved reasoning and generation quality
1y
100K+
21
Gemma 4: multimodal open AI models by Google, optimized for reasoning, coding, and long context.
6m
50K+
24B multimodal instruction model by Mistral AI, tuned for accuracy, tool use & fewer repeats
12m
10K+
1
SmolLM3 is a 3.1B model for efficient on-device use, with strong performance in chat
1y
50K+
9
Kimi K2 Thinking: open-source agent with deep reasoning, stable tool use, fast INT4, 256k context.
10m
50K+
4
Image generation model, uses a base latent diffusion model plus a refiner.
8m
50K+
8
Ministral 3: compact vision-enabled model with near-24B performance, optimized for local edge use
10m
50K+
4
Google’s latest Gemma, in its QAT (quantization aware trained) variant
12m
100K+
23
Granite Docling is a multimodal model for efficient document conversion.
11m
50K+
2
Efficient 284B MoE language model with 1M token context and multi-mode reasoning capabilities
5m
50K+
3
1T MoE multimodal agentic model with long-horizon coding, swarm orchestration, and native vision
5m
10K+
1