Community Organization
ELK-AI
Displaying 1 to 10 of 10 repositories
🦌 Qwen3-VL-32B NVFP4 | Vision-Language | 20GB (was 62GB) | <0.3% accuracy loss
9m
4.1K
1
Nemotron3-30B NVFP4 Quantized - First NVFP4 for NVIDIA Nemotron-3
9m
10K+
Zero-Config TensorRT-LLM | Pre-loaded Nemotron-3-Nano-30B | By Mutaz Al Awamleh - ELK-AI
9m
510
Devstral-Small-2-24B FP8 - Blackwell-optimized Mistral coding model
9m
423
NVIDIA Nemotron-VL-12B NVFP4 - Blackwell-optimized multimodal vision-language model
9m
1.4K
Qwen3-VL-2B-Thinking NVFP4 W4A16 - First NVFP4 quantization for Blackwell
9m
415
Qwen3-VL-4B-Thinking NVFP4 W4A16 - First NVFP4 quantization for Blackwell
9m
351
Qwen3-VL-8B-Thinking NVFP4 W4A16 - First NVFP4 quantization for Blackwell
9m
359
vLLM NVFP4 CUDA 13 - Blackwell-optimized vLLM for NVFP4/FP8 serving
10m
2.6K