nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4 Text Generation • 303B • Updated 11 days ago • 441k • • 321
RedHatAI/Qwen3-30B-A3B-Instruct-2507-quantized.w4a16 Text Generation • 5B • Updated Mar 20 • 5.2k • 3
RedHatAI/Mistral-Small-24B-Instruct-2501-quantized.w4a16 Text Generation • 24B • Updated 30 days ago • 5.28k • 2
Running 16 The Gemma Challenge and the Case for Agent Collabs 📝 16 Running an open agent collaboration to speed up Gemma 4
Running 31 Weight-Space Geometry of Offline Reasoning Training 🧭 31 Interactive weight-space geometry of six reasoning losses