Inference Providers
Active filters: vLLM
mistralai/Mistral-Medium-3.5-128B
128B • Updated • 166k
• 411
QuantTrio/GLM-5.2-Int4-Int8Mix
Text Generation
• 785B • Updated • 28.7k
• 14
QuantTrio/Qwen3.6-35B-A3B-AWQ
Image-Text-to-Text
• 36B • Updated • 1.17M
• 33
mistralai/Mistral-Small-4-119B-2603-eagle
Updated • 1.86k
• 56
unsloth/Mistral-Small-4-119B-2603-GGUF
119B • Updated • 14.1k
• 83
QuantTrio/Qwen3-Coder-30B-A3B-Instruct-AWQ
Text Generation
• 31B • Updated • 236k
• 9
QuantTrio/Qwen3-VL-30B-A3B-Instruct-AWQ
Text Generation
• 31B • Updated • 1.24M
• 44
mistralai/Mistral-Small-4-119B-2603
119B • Updated • 191k
• 411
mistralai/Mistral-Small-4-119B-2603-NVFP4
Updated • 2.61k
• 110
QuantTrio/gemma-4-31B-it-AWQ-6Bit
Image-Text-to-Text
• 31B • Updated • 1.91k
• 11
QuantTrio/Qwen3.6-27B-AWQ
Image-Text-to-Text
• 28B • Updated • 800k
• 23
QuantTrio/Qwen3.6-27B-AWQ-6Bit
Image-Text-to-Text
• 28B • Updated • 138k
• 18
mistralai/Mistral-Medium-3.5-128B-EAGLE
Updated • 212
• 57
festr2/GLM-5.2-Int8Mix-NVFP4
Text Generation
• Updated • 88
• 3
RedHatAI/DeepSeek-V4-Pro-NVFP4-FP8
Text Generation
• 894B • Updated • 384
• 1
Text Generation
• 173B • Updated • 1.93k
• 3
model-scope/glm-4-9b-chat-GPTQ-Int4
Text Generation
• 9B • Updated • 181
• 6
model-scope/glm-4-9b-chat-GPTQ-Int8
Text Generation
• 9B • Updated • 3
• 2
tclf90/qwen2.5-72b-instruct-gptq-int4
Text Generation
• 73B • Updated • 81
• 2
tclf90/qwen2.5-72b-instruct-gptq-int3
Text Generation
• 69B • Updated • 88
prithivMLmods/Nu2-Lupi-Qwen-14B
Text Generation
• 15B • Updated • 7
• 2
mradermacher/Nu2-Lupi-Qwen-14B-GGUF
15B • Updated • 69
• 1
mradermacher/Nu2-Lupi-Qwen-14B-i1-GGUF
15B • Updated • 538
• 1
JunHowie/Qwen3-0.6B-GPTQ-Int4
Text Generation
• 0.6B • Updated • 525
• 1
JunHowie/Qwen3-0.6B-GPTQ-Int8
Text Generation
• 0.6B • Updated • 6
JunHowie/Qwen3-1.7B-GPTQ-Int4
Text Generation
• 2B • Updated • 8.08k
• 1
JunHowie/Qwen3-1.7B-GPTQ-Int8
Text Generation
• 2B • Updated • 9
JunHowie/Qwen3-32B-GPTQ-Int4
Text Generation
• 33B • Updated • 9.2k
• 4
JunHowie/Qwen3-32B-GPTQ-Int8
Text Generation
• 33B • Updated • 254
• 4
JunHowie/Qwen3-30B-A3B-GPTQ-Int4
Text Generation
• 5B • Updated • 72
• 1