Inference Providers
Active filters: gptq
Text Generation
• 299B • Updated • 442
• 6
Qwen/Qwen3.5-122B-A10B-GPTQ-Int4
Image-Text-to-Text
• 125B • Updated • 259k
• 45
palmfuture/Qwen3.6-35B-A3B-GPTQ-Int4
Image-Text-to-Text
• 36B • Updated • 129k
• 28
canada-quant/GLM-5.2-W4A16-MTP
Text Generation
• 116B • Updated • 9.42k
• 19
compute1/Agents-A1-GPTQ-INT4-Sym
Text Generation
• 7B • Updated • 473
• 2
canada-quant/hy3-w4a16-mtp
Text Generation
• 47B • Updated • 578
• 2
Qwen/Qwen2-1.5B-Instruct-GPTQ-Int4
Text Generation
• 2B • Updated • 36.9k
• 6
tencent/Hunyuan-A13B-Instruct-GPTQ-Int4
Text Generation
• 80B • Updated • 257
• 53
baichuan-inc/Baichuan-M3-235B-GPTQ-INT4
Text Generation
• 235B • Updated • 560
• 11
Qwen/Qwen3.5-27B-GPTQ-Int4
Image-Text-to-Text
• 28B • Updated • 68.7k
• 57
Qwen/Qwen3.5-35B-A3B-GPTQ-Int4
Image-Text-to-Text
• 36B • Updated • 681k
• 93
canada-quant/DeepSeek-V4-Flash-W4A16-FP8
Text Generation
• 44B • Updated • 913
• 17
LordNeel/DeepSeek-V4-Flash-Acti-MTP-W4A16-FP8
Text Generation
• 44B • Updated • 3.89k
• 16
canada-quant/DeepSeek-V4-Flash-W4A16-FP8-MTP
Text Generation
• 51B • Updated • 3.54k
• 19
TheStageAI/gemma-4-E2B-it
Image-Text-to-Text
• Updated • 150
• 5
TheStageAI/gemma-4-E4B-it
Image-Text-to-Text
• Updated • 134
• 10
DuoNeural/Cosmos3-Nano-GPTQ-4bit-Abliterated
Text-to-Video
• Updated • 77
• 2
Sebesky/MiniMax-M3-W4A16-GPTQ
Image-Text-to-Text
• 430B • Updated • 4.44k
• 3
prasannaJagadesh/marlin-2B-GPTQ-4BITS
Video-Text-to-Text
• 2B • Updated • 171
• 3
elinas/alpaca-13b-lora-int4
Text Generation
• Updated • 19
• 40
elinas/alpaca-30b-lora-int4
Text Generation
• Updated • 16
• 68
mayaeary/pygmalion-6b-4bit-128g
Text Generation
• Updated • 12
• 40
mayaeary/pygmalion-6b_dev-4bit-128g
Text Generation
• Updated • 11
• 121
mayaeary/PPO_Pygway-V8p4_Dev-6b-4bit-128g
Text Generation
• Updated • 6
• 2
mayaeary/PPO_Pygway-6b-Mix-4bit-128g
Text Generation
• Updated • 3
• 2
Text Generation
• Updated • 11
• 45
Text Generation
• 7B • Updated • 61
• 31
Text Generation
• Updated • 144
• • 20
Text Generation
• Updated • 158
• 40
Text Generation
• 13B • Updated • 54
• 38