Sign inSign up

ELK-AI

Community Organization

ELK-AI

Displaying 1 to 10 of 10 repositories

image

🦌 Qwen3-VL-32B NVFP4 | Vision-Language | 20GB (was 62GB) | <0.3% accuracy loss

9m

4.1K

1

image

Nemotron3-30B NVFP4 Quantized - First NVFP4 for NVIDIA Nemotron-3

9m

10K+

image

Zero-Config TensorRT-LLM | Pre-loaded Nemotron-3-Nano-30B | By Mutaz Al Awamleh - ELK-AI

9m

510

image

Devstral-Small-2-24B FP8 - Blackwell-optimized Mistral coding model

9m

423

image

NVIDIA Nemotron-VL-12B NVFP4 - Blackwell-optimized multimodal vision-language model

9m

1.4K

image

Qwen3-VL-2B-Thinking NVFP4 W4A16 - First NVFP4 quantization for Blackwell

9m

415

image

Qwen3-VL-4B-Thinking NVFP4 W4A16 - First NVFP4 quantization for Blackwell

9m

351

image

Qwen3-VL-8B-Thinking NVFP4 W4A16 - First NVFP4 quantization for Blackwell

9m

359

image

vLLM NVFP4 CUDA 13 - Blackwell-optimized vLLM for NVFP4/FP8 serving

10m

2.6K