Twu31/Qwen3.6-35B-A3B-GPTQ-INT4-W4A16-LowLatency Image-Text-to-Text • 6B • Updated 6 days ago • 287 • 2
RedHatAI/DeepSeek-R1-Distill-Qwen-32B-quantized.w4a16 Text Generation • 33B • Updated 20 days ago • 2.54k • 6
RedHatAI/Mistral-Small-3.1-24B-Instruct-2503-quantized.w4a16 Image-Text-to-Text • 24B • Updated 12 days ago • 3.68k • 11
RedHatAI/Qwen3-Next-80B-A3B-Instruct-quantized.w4a16 Text Generation • 12B • Updated Apr 28 • 11.6k • 4
AEON-7/Qwen3.6-27B-AEON-Ultimate-Uncensored-NVFP4 Text Generation • 19B • Updated Jul 15 • 6.79k • 88
RedHatAI/Meta-Llama-3.1-8B-Instruct-quantized.w8a8 Text Generation • 8B • Updated 5 days ago • 28.5k • 20
RedHatAI/Meta-Llama-3.1-8B-Instruct-quantized.w4a16 Text Generation • 8B • Updated Jul 10 • 99.1k • 30
RedHatAI/Llama-3.2-3B-Instruct-quantized.w8a8 Text Generation • 4B • Updated Jul 10, 2025 • 1.49k • 1