Inference Providers
Active filters: cuda
Text-to-Video
• Updated • 1
alimpfard/breeze-tts-2-int4
Text-to-Speech
• Updated • 40
• 1
Barding-Defense/Qwen3.8-27B-huihui-abliterated-groupwise-int-NInfer
Image-Text-to-Text
• Updated • 7.84k
• 1
ruwwww/ornith-1.5-9b-ninfer
Text Generation
• Updated • 163
• 1
cometkim/Qwen3.8-27B-nvfp4qat-NInfer
Image-Text-to-Text
• Updated • 335
• 3
YousefAhmad121/Spark-X2.5-4B-Q8_0-GGUF
Text Generation
• 4B • Updated • 155
• 1
TechnoBaptist/Ternary-Bonsai-2-27B-mlx-2bit
Text Generation
• 27B • Updated • 1.68k
• 1
NoxNotreve/Ternary-Bonsai-2-27B-gguf
Text Generation
• 27B • Updated • 779
• 1
nuottroisaoduoc/Bonsai-2-27B-Ternary-CRACK-GGUF
Text Generation
• 27B • Updated • 637
• 1
Text Generation
• 6B • Updated • 1
Text Generation
• 0.8B • Updated • 1
Text Generation
• 4B • Updated • 1
JMVRoill/Qwen3.8-27B-Uncensored-nvfp4-NInfer
Image-Text-to-Text
• Updated • 374
• 3
Text Generation
• Updated • 75
• 23
CalderaAI/13B-Ouroboros-GPTQ4bit-128g-CUDA
Text Generation
• Updated • 11
marcorez8/llama-cpp-python-windows-blackwell-cuda
ValiantLabs/Qwen3-8B-ShiningValiant3
Text Generation
• 8B • Updated • 28
• 3
mradermacher/Qwen3-8B-ShiningValiant3-GGUF
8B • Updated • 1.53k
• 2
mradermacher/Qwen3-8B-ShiningValiant3-i1-GGUF
8B • Updated • 1.29k
• 2
ValiantLabs/Qwen3-1.7B-ShiningValiant3
Text Generation
• 2B • Updated • 25
• 5
mradermacher/Qwen3-1.7B-ShiningValiant3-GGUF
2B • Updated • 461
mradermacher/Qwen3-1.7B-ShiningValiant3-i1-GGUF
2B • Updated • 773
ValiantLabs/Qwen3-4B-ShiningValiant3
Text Generation
• 4B • Updated • 25
• 7
sequelbox/Qwen3-8B-PlumEsper
Text Generation
• 8B • Updated • 98
• 1
sequelbox/Qwen3-4B-PlumEsper
Text Generation
• 4B • Updated • 24
mradermacher/Qwen3-Shining-Lucy-CODER-3.5B-Brainstorm20x-e32-GGUF
3B • Updated • 520
• 1
mradermacher/Qwen3-Shining-Lucy-CODER-2.4B-mix2-GGUF
2B • Updated • 542
mradermacher/Qwen3-Shining-Lucy-CODER-2.4B-GGUF
2B • Updated • 532
mradermacher/Qwen3-Shining-Lucy-CODER-2.4B-mix2-i1-GGUF
2B • Updated • 1.03k