WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation Paper • 2608.24479 • Published 6 days ago • 135
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published 19 days ago • 289
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 21 days ago • 342
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published Jul 30 • 309
FVAttn: Adaptive Sparse Attention with Runtime Load Balancing for Video Generation Paper • 2607.16190 • Published Jul 17 • 9
Self in Space: Benchmarking Self-Awareness and Spatial Cognition in UAV Embodied Intelligence Paper • 2607.12477 • Published Jul 14 • 11
SAM-MT: Real-Time Interactive Multi-Target Video Segmentation Paper • 2607.08688 • Published Jul 9 • 9
Mean Mode Screaming: Mean--Variance Split Residuals for 1000-Layer Diffusion Transformers Paper • 2605.06169 • Published May 7 • 238
DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards Paper • 2605.21467 • Published May 20 • 207
SpecBench: Measuring Reward Hacking in Long-Horizon Coding Agents Paper • 2605.21384 • Published May 20 • 4