Inference Providers
Active filters: vLLM
mistralai/Mistral-Medium-3.5-128B
128B • Updated • 310k
• 389
unsloth/Mistral-Small-4-119B-2603-GGUF
119B • Updated • 9.3k
• 80
mistralai/Mistral-Small-4-119B-2603
119B • Updated • 153k
• 404
QuantTrio/Qwen3.6-27B-AWQ-6Bit
Image-Text-to-Text
• 28B • Updated • 53.2k
• 15
QuantTrio/Qwen3.6-27B-AWQ
Image-Text-to-Text
• 28B • Updated • 1.4M
• 19
QuantTrio/GLM-5.2-Int4-Int8Mix
Text Generation
• 785B • Updated • 59.3k
• 7
Text Generation
• 426B • Updated • 193
• 2
QuantTrio/gemma-4-31B-it-AWQ
Image-Text-to-Text
• 31B • Updated • 502k
• 13
QuantTrio/Qwen3.6-35B-A3B-AWQ
Image-Text-to-Text
• 36B • Updated • 722k
• 27
mistralai/Mistral-Medium-3.5-128B-EAGLE
Updated • 299
• 50
model-scope/glm-4-9b-chat-GPTQ-Int4
Text Generation
• 9B • Updated • 118
• 6
model-scope/glm-4-9b-chat-GPTQ-Int8
Text Generation
• 9B • Updated • 12
• 2
tclf90/qwen2.5-72b-instruct-gptq-int4
Text Generation
• 73B • Updated • 100
• 2
tclf90/qwen2.5-72b-instruct-gptq-int3
Text Generation
• 69B • Updated • 165
prithivMLmods/Nu2-Lupi-Qwen-14B
Text Generation
• 15B • Updated • 5
• 2
mradermacher/Nu2-Lupi-Qwen-14B-GGUF
15B • Updated • 71
• 1
mradermacher/Nu2-Lupi-Qwen-14B-i1-GGUF
15B • Updated • 210
• 1
JunHowie/Qwen3-0.6B-GPTQ-Int4
Text Generation
• 0.6B • Updated • 199
• 1
JunHowie/Qwen3-0.6B-GPTQ-Int8
Text Generation
• 0.6B • Updated • 6
JunHowie/Qwen3-1.7B-GPTQ-Int4
Text Generation
• 2B • Updated • 2.13k
• 1
JunHowie/Qwen3-1.7B-GPTQ-Int8
Text Generation
• 2B • Updated • 8
JunHowie/Qwen3-32B-GPTQ-Int4
Text Generation
• 33B • Updated • 1.8k
• 4
JunHowie/Qwen3-32B-GPTQ-Int8
Text Generation
• 33B • Updated • 149
• 4
JunHowie/Qwen3-30B-A3B-GPTQ-Int4
Text Generation
• 5B • Updated • 74
• 1
JunHowie/Qwen3-14B-GPTQ-Int8
Text Generation
• 15B • Updated • 96
• 1
JunHowie/Qwen3-14B-GPTQ-Int4
Text Generation
• 15B • Updated • 170k
• 4
JunHowie/Qwen3-8B-GPTQ-Int8
Text Generation
• 8B • Updated • 2.19k
JunHowie/Qwen3-8B-GPTQ-Int4
Text Generation
• 8B • Updated • 369
• 4
JunHowie/Qwen3-4B-GPTQ-Int4
Text Generation
• 4B • Updated • 933
• 1
JunHowie/Qwen3-4B-GPTQ-Int8
Text Generation
• 4B • Updated • 18