Inference Providers
Active filters: w4a16
RedHatAI/gemma-3-1b-it-quantized.w4a16
Text Generation
• 1B • Updated • 23
adamrb/mpt-30b-chat-w4a16-gptq
30B • Updated • 3
RedHatAI/SmolLM3-3B-quantized.w4a16
3B • Updated • 19
• 1
ramblingpolymath/Qwen3-30B-A3B-Instruct-2507-W4A16
Text Generation
• 31B • Updated • 10
ramblingpolymath/Qwen3-Coder-30B-A3B-Instruct-W4A16
Text Generation
• 31B • Updated • 29
• 2
ramblingpolymath/Qwen3-30B-A3B-Thinking-2507-W4A16
Text Generation
• 31B • Updated • 5
• 3
twhitworth/gpt-oss-120b-awq-w4a16
117B • Updated • 3.29k
• 24
TheHouseOfTheDude/Behemoth-R1-123B-v2_Compressed-Tensors
Text Generation
• Updated • 2
TheHouseOfTheDude/Behemoth-X-123B-v2_Compressed-Tensors
Text Generation
• Updated TheHouseOfTheDude/GLM-Steam-106B-A12B-v1_Compressed-Tensors
Text Generation
• Updated TheHouseOfTheDude/L3.3-Animus-V10.0_Compressed-Tensors
Text Generation
• Updated TheHouseOfTheDude/Behemoth-ReduX-123B-v1_Compressed-Tensors
Text Generation
• Updated • 3
TheHouseOfTheDude/Qwen3-Next-80B-A3B-Instruct_Compressed-Tensors
Text Generation
• Updated • 8
TheHouseOfTheDude/Fallen-Command-A-111B-v1_Compresses-Tensors
Text Generation
• Updated TheHouseOfTheDude/Behemoth-ReduX-123B-v1.1_Compressed-Tensors
Text Generation
• Updated • 4
TheHouseOfTheDude/L3.3-70B-Animus-V12.0_Compressed-Tensors
Text Generation
• Updated TheHouseOfTheDude/Behemoth-X-123B-v2.1_Compressed-Tensors
Text Generation
• Updated • 1
ModelCloud/GLM-4.6-GPTQMODEL-W4A16-v1
Text Generation
• 357B • Updated • 6
ModelCloud/GLM-4.6-GPTQMODEL-W4A16-v2
Text Generation
• 357B • Updated • 13
• 1
ModelCloud/GLM-4.6-REAP-268B-A32B-GPTQMODEL-W4A16
Text Generation
• 269B • Updated • 1
• 2
ModelCloud/MiniMax-M2-GPTQMODEL-W4A16
Text Generation
• 229B • Updated • 13
• 3
ModelCloud/Marin-32B-Base-GPTQMODEL-W4A16
Text Generation
• 33B • Updated • 5
• 1
ModelCloud/Marin-32B-Base-GPTQMODEL-AWQ-W4A16
Text Generation
• 33B • Updated • 9
• 2
TheHouseOfTheDude/Legion-V2.1-LLaMa-70B_CompressedTensors
Text Generation
• Updated ModelCloud/Granite-4.0-H-1B-GPTQMODEL-W4A16
Text Generation
• 1B • Updated • 6
• 1
ModelCloud/Granite-4.0-H-350M-GPTQMODEL-W4A16
Text Generation
• 0.3B • Updated • 4
• 1
ModelCloud/Brumby-14B-Base-GPTQMODEL-W4A16
Text Generation
• 15B • Updated • 6
• 1
ModelCloud/Brumby-14B-Base-GPTQMODEL-W4A16-v2
Text Generation
• 15B • Updated • 5
• 1
TheHouseOfTheDude/Precog-24B-v1_Compressed-Tensors
Text Generation
• Updated • 2
TheHouseOfTheDude/M2411-123B-Animus-V12.0_Compressed-Tensors
Text Generation
• Updated • 2