E-MoE: Enhanced Mixture-of-Experts for Non-Factorized Diffusion Language Models Paper • 2609.37533 • Published 13 days ago • 67
Decentralized Master-Mind: Joint Action Refinement through Iterative Intent Denoising in Multi-Agent Pathfinding Paper • 2609.32019 • Published 17 days ago • 64
LANTERN: Illuminating Hidden Mathematical Knowledge in Language Models Paper • 2609.32264 • Published 16 days ago • 50
Your Transformer Can Hold Two Thoughts at Once: Evidence of Linear Superposition in LLMs Paper • 2609.29845 • Published 18 days ago • 105
Gemma 4 AmbigQA — OFT Collection OFT-adapted Gemma 4 E2B and E4B ensemble members trained on AmbigQA. • 15 items • Updated Aug 7
Gemma 4 AmbigQA — OFT Collection OFT-adapted Gemma 4 E2B and E4B ensemble members trained on AmbigQA. • 15 items • Updated Aug 7
Gemma 4 AmbigQA — OFT Collection OFT-adapted Gemma 4 E2B and E4B ensemble members trained on AmbigQA. • 15 items • Updated Aug 7
Gemma 4 AmbigQA — OFT Collection OFT-adapted Gemma 4 E2B and E4B ensemble members trained on AmbigQA. • 15 items • Updated Aug 7
Gemma 4 AmbigQA — OFT Collection OFT-adapted Gemma 4 E2B and E4B ensemble members trained on AmbigQA. • 15 items • Updated Aug 7