Inference Providers
Active filters: vLLM
mistralai/Mistral-Small-4-119B-2603-NVFP4
Updated • 1.35k
• 120
mistralai/Mistral-Medium-3.5-128B
128B • Updated • 144k
• 439
mistralai/Mistral-Small-4-119B-2603
119B • Updated • 55.7k
• 424
Image-Text-to-Text
• 10B • Updated • 1.14M
• 28
QuantTrio/GLM-4.5-Air-GPTQ-Int4-Int8Mix
Text Generation
• 20B • Updated • 497
• 11
mistralai/Mistral-Small-4-119B-2603-eagle
Updated • 307
• 59
bartowski/mistralai_Mistral-Small-4-119B-2603-GGUF
Image-Text-to-Text
• 119B • Updated • 3.44k
• 17
unsloth/Mistral-Small-4-119B-2603-GGUF
119B • Updated • 12.3k
• 85
QuantTrio/Qwen3.6-35B-A3B-AWQ
Image-Text-to-Text
• 36B • Updated • 596k
• 35
mistralai/Mistral-Medium-3.5-128B-EAGLE
Updated • 311
• 58
model-scope/glm-4-9b-chat-GPTQ-Int4
Text Generation
• 9B • Updated • 75
• 6
model-scope/glm-4-9b-chat-GPTQ-Int8
Text Generation
• 9B • Updated • 24
• 2
tclf90/qwen2.5-72b-instruct-gptq-int4
Text Generation
• 73B • Updated • 71
• 2
tclf90/qwen2.5-72b-instruct-gptq-int3
Text Generation
• 69B • Updated • 61
prithivMLmods/Nu2-Lupi-Qwen-14B
Text Generation
• 15B • Updated • 7
• 2
mradermacher/Nu2-Lupi-Qwen-14B-GGUF
15B • Updated • 529
• 1
mradermacher/Nu2-Lupi-Qwen-14B-i1-GGUF
15B • Updated • 875
• 1
JunHowie/Qwen3-0.6B-GPTQ-Int4
Text Generation
• 0.6B • Updated • 582
• 1
JunHowie/Qwen3-0.6B-GPTQ-Int8
Text Generation
• 0.6B • Updated • 9
JunHowie/Qwen3-1.7B-GPTQ-Int4
Text Generation
• 2B • Updated • 986
• 1
JunHowie/Qwen3-1.7B-GPTQ-Int8
Text Generation
• 2B • Updated • 299
JunHowie/Qwen3-32B-GPTQ-Int4
Text Generation
• 33B • Updated • 1.46k
• 4
JunHowie/Qwen3-32B-GPTQ-Int8
Text Generation
• 33B • Updated • 422
• 4
JunHowie/Qwen3-30B-A3B-GPTQ-Int4
Text Generation
• 5B • Updated • 43
• 1
JunHowie/Qwen3-14B-GPTQ-Int8
Text Generation
• 15B • Updated • 269
• 1
JunHowie/Qwen3-14B-GPTQ-Int4
Text Generation
• 15B • Updated • 1.01k
• 4
JunHowie/Qwen3-8B-GPTQ-Int8
Text Generation
• 8B • Updated • 3.18k
JunHowie/Qwen3-8B-GPTQ-Int4
Text Generation
• 8B • Updated • 1.88k
• 4
JunHowie/Qwen3-4B-GPTQ-Int4
Text Generation
• 4B • Updated • 735
• 1
JunHowie/Qwen3-4B-GPTQ-Int8
Text Generation
• 4B • Updated • 332