Inference Providers
Active filters: sglang
Soomin33/Qwen3.8-Flash-Next-FP6-INT8
Text Generation
• 98B • Updated • 199
• 4
HaberstrohSystems/Qwen3.8-Flash-Next-int2-mixed-AutoRound-24GB-SGLang
Text Generation
• 11B • Updated • 635
• 7
Anbeeld/Ling-3.0-flash-DSpark-GGUF
Text Generation
• 1.0B • Updated • 702
• 1
tencent/Simple-Attention-Sparsification
Text Generation
• Updated • 8
Text Generation
• 18B • Updated • 242
• 1
igor255/Qwen3.8-27B-DFlash2-EXL3-4.00bpw
Text Generation
• 0.6B • Updated • 195
• 1
cbert33/Agnes-3.0-Flash-FP8-Calibrated
Image-Text-to-Text
• 33B • Updated • 17
• 1
mradermacher/Qwen3.8-27B-Uncensored-Mythos-Class-Agentic-i1-GGUF
27B • Updated • 2.12k
• 1
superagent-ai/security-one-27b
Text Classification
• 27B • Updated • 1
kaushikvira/Qwen3.8-27B-thinkingcap-NVFP4-HF
Image-Text-to-Text
• 15B • Updated • 11
• 1
Blackfrost-AI/MiMo-V2.6-Flash-MOPD-Derisking-Intervention-Runtime
Text Generation
• Updated • 1
voska/GLM-5.3-Flash-Uncensored-W4A16
Image-Text-to-Text
• 321B • Updated • 781
• 4
SurfaceData/llava-v1.6-mistral-7b-sglang
Image-Text-to-Text
• 8B • Updated • 11
• 9
SurfaceData/llava-v1.6-vicuna-7b-sglang
Image-Text-to-Text
• 7B • Updated • 11
• 1
tclf90/qwen2.5-72b-instruct-gptq-int4
Text Generation
• 73B • Updated • 68
• 2
tclf90/qwen2.5-72b-instruct-gptq-int3
Text Generation
• 69B • Updated • 71
alvarobartt/grok-2-tokenizer
Text Generation
• Updated • 43
• 5
lemuralabs/MiniMax-M2-Pruned-25
173B • Updated • 464
• 36
mradermacher/MiniMax-M2-THRIFT-GGUF
JasmineBBB/Kimi-Linear-48B-A3B-Instruct-bnb-4bit
Text Generation
• 49B • Updated • 39
• 1
mradermacher/MiniMax-M2-THRIFT-i1-GGUF
173B • Updated • 365
• 10
bartowski/VibeStudio_MiniMax-M2-THRIFT-GGUF
Text Generation
• 173B • Updated • 1.62k
• 9
lemuralabs/MiniMax-M2-Pruned-55
106B • Updated • 89
• 5
JinnP/SGLang-EAGLE3-Qwen3-Coder-30B-A3B-Instruct
Text Generation
• 0.2B • Updated • 172
• 2
mradermacher/MiniMax-M2-THRIFT-55-GGUF
106B • Updated • 132
• 2
mradermacher/MiniMax-M2-THRIFT-55-i1-GGUF
106B • Updated • 1.05k
• 2
Doradus-AI/MiroThinker-v1.0-30B-FP8
Text Generation
• 31B • Updated • 30
• 4
Doradus-AI/Hermes-4.3-36B-FP8
Text Generation
• 36B • Updated • 534
• 5
Doradus-AI/RnJ-1-Instruct-FP8
Text Generation
• 9B • Updated • 94
• 5