Inference Providers
Active filters: vLLM
QuantTrio/DeepSeek-R1-0528-Qwen3-8B-GPTQ-Int4-Int8Mix
Text Generation
• 11B • Updated • 44
• 4
QuantTrio/DeepSeek-R1-0528-GPTQ-Int4-Int8Mix-Lite
Text Generation
• 721B • Updated • 51
• 2
QuantTrio/DeepSeek-R1-0528-GPTQ-Int4-Int8Mix-Compact
Text Generation
• 847B • Updated • 31
• 5
QuantTrio/DeepSeek-R1-0528-GPTQ-Int4-Int8Mix-Medium
Text Generation
• 912B • Updated • 462
• 1
brandonbeiler/InternVL3-38B-FP8-Dynamic
Image-Text-to-Text
• 38B • Updated • 11
brandonbeiler/InternVL3-78B-FP8-Dynamic
Image-Text-to-Text
• 78B • Updated • 54
brandonbeiler/InternVL3-8B-FP8-Dynamic
Image-Text-to-Text
• 8B • Updated • 16
• 2
dengcao/GLM-4.1V-9B-Thinking-GPTQ-Int4-Int8Mix
Image-Text-to-Text
• 15B • Updated • 36
• 2
dengcao/GLM-4.1V-9B-Thinking-AWQ
Image-Text-to-Text
• 10B • Updated • 47
• 1
brandonbeiler/Skywork-R1V3-38B-FP8-Dynamic
Image-Text-to-Text
• 38B • Updated • 24
• 2
QuantTrio/Qwen3-235B-A22B-Instruct-2507-GPTQ-Int4-Int8Mix
Text Generation
• 248B • Updated • 121
• 4
QuantTrio/Qwen3-235B-A22B-Instruct-2507-AWQ
Text Generation
• 235B • Updated • 8.01k
• 13
QuantTrio/GLM-4.1V-9B-Thinking-GPTQ-Int4-Int8Mix
Text Generation
• 15B • Updated • 30
• 1
QuantTrio/GLM-4.1V-9B-Thinking-AWQ
Text Generation
• 10B • Updated • 3.36k
QuantTrio/Qwen3-Coder-480B-A35B-Instruct-AWQ
Text Generation
• 480B • Updated • 247
• 8
QuantTrio/Qwen3-Coder-480B-A35B-Instruct-GPTQ-Int4-Int8Mix
Text Generation
• 534B • Updated • 61
• 7
QuantTrio/Qwen3-235B-A22B-Thinking-2507-AWQ
Text Generation
• 235B • Updated • 352
• 6
QuantTrio/Qwen3-235B-A22B-Thinking-2507-GPTQ-Int4-Int8Mix
Text Generation
• 253B • Updated • 47
• 4
QuantTrio/GLM-4.5-Air-AWQ-FP16Mix
Text Generation
• 24B • Updated • 2.13k
• 14
QuantTrio/GLM-4.5-Air-GPTQ-Int4-Int8Mix
Text Generation
• 20B • Updated • 257
• 11
QuantTrio/Qwen3-30B-A3B-Instruct-2507-GPTQ-Int8
Text Generation
• 31B • Updated • 299
• 9
QuantTrio/GLM-4.5-GPTQ-Int4-Int8Mix
Text Generation
• 55B • Updated • 41
• 5
Text Generation
• 53B • Updated • 72
• 9
QuantTrio/Qwen3-30B-A3B-Thinking-2507-AWQ-BF16Mix
Text Generation
• 31B • Updated • 84
• 5
QuantTrio/Qwen3-30B-A3B-Thinking-2507-GPTQ-Int8
Text Generation
• 31B • Updated • 63
• 2
QuantTrio/Qwen3-30B-A3B-Thinking-2507-AWQ
Text Generation
• 31B • Updated • 4.28k
• 4
QuantTrio/KAT-V1-40B-GPTQ-Int4-Int8Mix
Text Generation
• 47B • Updated • 14
QuantTrio/Qwen3-Coder-30B-A3B-Instruct-GPTQ-Int8
Text Generation
• 31B • Updated • 152k
• 8
QuantTrio/Qwen3-Coder-30B-A3B-Instruct-AWQ
Text Generation
• 31B • Updated • 311k
• 11
EliovpAI/Qwen3-14B-FP8-KV
Text Generation
• 15B • Updated • 24
• 2