Sign inSign up
1 - 30 of 109 results for ai/.
​
model + 1 more

30d

1M+

49

model + 1 more

2m

500K+

215

model + 1 more

2m

500K+

39

model + 1 more

30d

500K+

10

model + 1 more

2m

500K+

46

model + 1 more

2m

1M+

68

model

Solid LLaMA 3 update, reliable for coding, chat, and Q&A tasks

1y

500K+

32

model

Distilled LLaMA by DeepSeek, fast and optimized for real-world tasks

1y

100K+

79

model + 1 more

1m

100K+

5

Artifact

24d

100K+

1

model + 1 more

2m

100K+

33

model

Newest LLama 3 release with improved reasoning and generation quality

1y

100K+

21

model

SmolLM3 is a 3.1B model for efficient on-device use, with strong performance in chat

1y

100K+

9

model + 1 more

2m

100K+

11

model + 1 more

2m

100K+

7

model

Kimi K2 Thinking: open-source agent with deep reasoning, stable tool use, fast INT4, 256k context.

10m

50K+

4

model

Image generation model, uses a base latent diffusion model plus a refiner.

8m

50K+

8

model + 1 more

2m

50K+

1

model

Versatile Qwen update with better language skills and wider support

1y

100K+

13

model + 1 more

2m

100K+

11

image

llmman with llama.cpp: server, server-cuda, server-cuda13, server-rocm, server-vulkan

11h

100K+

model + 1 more

2m

50K+

3

model

Ministral 3: compact vision-enabled model with near-24B performance, optimized for local edge use

10m

50K+

4

model

Granite Docling is a multimodal model for efficient document conversion.

12m

50K+

2

Artifact

Efficient 284B MoE language model with 1M token context and multi-mode reasoning capabilities

5m

50K+

3

model

Google’s latest Gemma, in its QAT (quantization aware trained) variant

1y

100K+

23

model

Microsoft’s compact model, surprisingly capable at reasoning and code

1y

100K+

26

model

1T MoE multimodal agentic model with long-horizon coding, swarm orchestration, and native vision

5m

50K+

1

model

Ministral 3: compact vision-enabled model with near-24B performance, optimized for local edge use

10m

100K+

5

model

Efficient 80B MoE coding model with 3B activated params, 256K context, and agentic capabilities

8m

100K+

3