Inference Providers
Active filters: vLLM
Image-Text-to-Text
• 17B • Updated • 606
• 19
QuantTrio/Seed-OSS-36B-Instruct-AWQ
Text Generation
• 36B • Updated • 515
• 8
QuantTrio/Seed-OSS-36B-Instruct-GPTQ-Int8
Text Generation
• 36B • Updated • 26
• 4
QuantTrio/Seed-OSS-36B-Instruct-GPTQ-Int4
Text Generation
• 36B • Updated • 33
• 5
QuantTrio/Seed-OSS-36B-Instruct-GPTQ-Int3
Text Generation
• 34B • Updated • 16
• 3
amakhov/tiny-random-llama
Text Generation
• 4.18M • Updated • 306
Text Generation
• 41B • Updated • 17
• 2
QuantTrio/DeepSeek-V3.1-AWQ
Text Generation
• 684B • Updated • 138
• 5
QuantTrio/DeepSeek-V3.1-AWQ-Fp16Mix
Text Generation
• 684B • Updated • 54
• 1
QuantTrio/DeepSeek-V3.1-AWQ-Lite
Text Generation
• 684B • Updated • 87
• 3
JunHowie/Qwen3-4B-Instruct-2507-GPTQ-Int4
Text Generation
• 4B • Updated • 1.94k
• 4
JunHowie/Qwen3-4B-Instruct-2507-GPTQ-Int8
Text Generation
• 4B • Updated • 280
JunHowie/Qwen3-4B-Thinking-2507-GPTQ-Int4
Text Generation
• 4B • Updated • 35
• 1
JunHowie/Qwen3-4B-Thinking-2507-GPTQ-Int8
Text Generation
• 4B • Updated • 8
• 2
JunHowie/Qwen3-30B-A3B-Instruct-2507-GPTQ-Int4
Text Generation
• 31B • Updated • 1.14k
JunHowie/Qwen3-30B-A3B-Instruct-2507-GPTQ-Int8
Text Generation
• 31B • Updated • 7
JunHowie/Qwen3-30B-A3B-Thinking-2507-GPTQ-Int4
Text Generation
• 31B • Updated • 8
JunHowie/Qwen2-7B-Instruct-GPTQ-Int4
Text Generation
• 8B • Updated • 6
JunHowie/Qwen2-7B-Instruct-GPTQ-Int8
Text Generation
• 8B • Updated • 7
EliovpAI/Deepseek-R1-0528-Qwen3-8B-FP8-KV
Text Generation
• 8B • Updated • 45
JunHowie/Qwen3-30B-A3B-Thinking-2507-GPTQ-Int8
Text Generation
• 31B • Updated • 7
JunHowie/Seed-OSS-36B-Instruct-GPTQ-Int4
Text Generation
• 36B • Updated • 12
JunHowie/Seed-OSS-36B-Instruct-GPTQ-Int8
Text Generation
• 36B • Updated • 13
QuantTrio/Qwen3-VL-235B-A22B-Instruct-FP8
Text Generation
• Updated • 114
QuantTrio/Qwen3-VL-235B-A22B-Thinking-AWQ
Text Generation
• 236B • Updated • 1.44k
• 8
QuantTrio/Qwen3-VL-235B-A22B-Thinking-FP8
Text Generation
• 236B • Updated • 64
QuantTrio/DeepSeek-V3.2-Exp-AWQ
Text Generation
• 685B • Updated • 1.78k
• 4
QuantTrio/DeepSeek-V3.2-Exp-AWQ-Lite
Text Generation
• 685B • Updated • 40
• 4
Text Generation
• 50B • Updated • 762
• 5
QuantTrio/GLM-4.6-GPTQ-Int4-Int8Mix
Text Generation
• 69B • Updated • 21
• 4