Inference Providers
Active filters: int4
study-hjt/Qwen1.5-110B-Chat-GPTQ-Int4
Text Generation
• 111B • Updated • 7
• 2
study-hjt/CodeQwen1.5-7B-Chat-GPTQ-Int4
Text Generation
• 7B • Updated • 6
study-hjt/Qwen1.5-110B-Chat-AWQ
Text Generation
• 111B • Updated • 5
modelscope/Yi-1.5-34B-Chat-AWQ
Text Generation
• 34B • Updated • 153
• 2
modelscope/Yi-1.5-6B-Chat-GPTQ
Text Generation
• 6B • Updated • 7
modelscope/Yi-1.5-6B-Chat-AWQ
Text Generation
• 6B • Updated • 8
modelscope/Yi-1.5-9B-Chat-GPTQ
Text Generation
• 9B • Updated • 9
• 1
modelscope/Yi-1.5-9B-Chat-AWQ
Text Generation
• 9B • Updated • 78
modelscope/Yi-1.5-34B-Chat-GPTQ
Text Generation
• 34B • Updated • 7
• 1
jojo1899/Phi-3-mini-128k-instruct-ov-int4
Text Generation
• Updated • 10
jojo1899/Llama-2-13b-chat-hf-ov-int4
Text Generation
• Updated • 7
jojo1899/Mistral-7B-Instruct-v0.2-ov-int4
Text Generation
• Updated • 13
model-scope/glm-4-9b-chat-GPTQ-Int4
Text Generation
• 9B • Updated • 143
• 6
ModelCloud/Mistral-Nemo-Instruct-2407-gptq-4bit
Text Generation
• 12B • Updated • 292
• 5
ModelCloud/Meta-Llama-3.1-8B-Instruct-gptq-4bit
Text Generation
• 8B • Updated • 2.99k
• 4
ModelCloud/Meta-Llama-3.1-8B-gptq-4bit
Text Generation
• 8B • Updated • 85
ModelCloud/Meta-Llama-3.1-70B-Instruct-gptq-4bit
Text Generation
• 71B • Updated • 10
• 4
ModelCloud/Mistral-Large-Instruct-2407-gptq-4bit
Text Generation
• 123B • Updated • 8
• 1
RedHatAI/Meta-Llama-3.1-8B-Instruct-quantized.w4a16
Text Generation
• 8B • Updated • 88.8k
• 30
angeloc1/llama3dot1SimilarProcesses4
Text Generation
• 8B • Updated • 5
angeloc1/llama3dot1DifferentProcesses4
Text Generation
• 8B • Updated • 4
ModelCloud/Meta-Llama-3.1-405B-Instruct-gptq-4bit
Text Generation
• 410B • Updated • 6
• 2
RedHatAI/Meta-Llama-3.1-70B-Instruct-quantized.w4a16
Text Generation
• 71B • Updated • 45.3k
• 33
ModelCloud/EXAONE-3.0-7.8B-Instruct-gptq-4bit
8B • Updated • 6
• 3
RedHatAI/Meta-Llama-3.1-405B-Instruct-quantized.w4a16
Text Generation
• 406B • Updated • 275
• 11
angeloc1/llama3dot1FoodDel4v05
Text Generation
• 8B • Updated • 4
zzzmahesh/Meta-Llama-3-8B-Instruct-quantized.w4a4
Text Generation
• 8B • Updated • 8
• 1
ModelCloud/GRIN-MoE-gptq-4bit
42B • Updated • 4
• 6
joshmiller656/Llama3.2-1B-AWQ-INT4
1B • Updated • 8
Advantech-EIOT/intel_llama-3.1-8b-instruct