Inference Providers
Active filters: vLLM
Image-Text-to-Text
• 17B • Updated • 415
• 19
QuantTrio/Seed-OSS-36B-Instruct-AWQ
Text Generation
• 36B • Updated • 386
• 8
QuantTrio/Seed-OSS-36B-Instruct-GPTQ-Int8
Text Generation
• 36B • Updated • 50
• 4
QuantTrio/Seed-OSS-36B-Instruct-GPTQ-Int4
Text Generation
• 36B • Updated • 41
• 5
QuantTrio/Seed-OSS-36B-Instruct-GPTQ-Int3
Text Generation
• 34B • Updated • 21
• 3
amakhov/tiny-random-llama
Text Generation
• 4.18M • Updated • 261
Text Generation
• 41B • Updated • 21
• 2
QuantTrio/DeepSeek-V3.1-AWQ
Text Generation
• 684B • Updated • 752
• 5
QuantTrio/DeepSeek-V3.1-AWQ-Fp16Mix
Text Generation
• 684B • Updated • 55
• 1
QuantTrio/DeepSeek-V3.1-AWQ-Lite
Text Generation
• 684B • Updated • 39
• 3
JunHowie/Qwen3-4B-Instruct-2507-GPTQ-Int4
Text Generation
• 4B • Updated • 1.67k
• 4
JunHowie/Qwen3-4B-Instruct-2507-GPTQ-Int8
Text Generation
• 4B • Updated • 182
JunHowie/Qwen3-4B-Thinking-2507-GPTQ-Int4
Text Generation
• 4B • Updated • 49
• 1
JunHowie/Qwen3-4B-Thinking-2507-GPTQ-Int8
Text Generation
• 4B • Updated • 11
• 2
JunHowie/Qwen3-30B-A3B-Instruct-2507-GPTQ-Int4
Text Generation
• 31B • Updated • 782
JunHowie/Qwen3-30B-A3B-Instruct-2507-GPTQ-Int8
Text Generation
• 31B • Updated • 21
JunHowie/Qwen3-30B-A3B-Thinking-2507-GPTQ-Int4
Text Generation
• 31B • Updated • 34
JunHowie/Qwen2-7B-Instruct-GPTQ-Int4
Text Generation
• 8B • Updated • 12
JunHowie/Qwen2-7B-Instruct-GPTQ-Int8
Text Generation
• 8B • Updated • 10
EliovpAI/Deepseek-R1-0528-Qwen3-8B-FP8-KV
Text Generation
• 8B • Updated • 44
JunHowie/Qwen3-30B-A3B-Thinking-2507-GPTQ-Int8
Text Generation
• 31B • Updated • 12
JunHowie/Seed-OSS-36B-Instruct-GPTQ-Int4
Text Generation
• 36B • Updated • 24
JunHowie/Seed-OSS-36B-Instruct-GPTQ-Int8
Text Generation
• 36B • Updated • 18
QuantTrio/Qwen3-VL-235B-A22B-Instruct-AWQ
Text Generation
• 236B • Updated • 16k
• 14
QuantTrio/Qwen3-VL-235B-A22B-Instruct-FP8
Text Generation
• Updated • 72
QuantTrio/Qwen3-VL-235B-A22B-Thinking-AWQ
Text Generation
• 236B • Updated • 951
• 8
QuantTrio/Qwen3-VL-235B-A22B-Thinking-FP8
Text Generation
• 236B • Updated • 22
QuantTrio/DeepSeek-V3.2-Exp-AWQ
Text Generation
• 685B • Updated • 1.85k
• 4
QuantTrio/DeepSeek-V3.2-Exp-AWQ-Lite
Text Generation
• 685B • Updated • 32
• 4
Text Generation
• 50B • Updated • 245
• 5