Business Arena: Benchmarking LLM Agents in a Realistic Marketplace Paper • 2608.08621 • Published 3 days ago • 3
The Optimizer Is the Agent: Reasoning-Driven Search across Prompts, Programs, and ML Workflows Paper • 2608.06714 • Published 5 days ago • 6
DuplexGen: Adaptive Synthesis of Human-AI Turn-Taking Dialogues Paper • 2607.26178 • Published 15 days ago • 8
Beyond Simply Environment Scaling: Designing Effective Environment Distributions for Multimodal Agent Learning Paper • 2608.03571 • Published 6 days ago • 40
SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs Paper • 2608.03573 • Published 6 days ago • 46
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published 8 days ago • 27
RynnValue: Scaling Robotic Value Foundation Models with Temporal Distance Paper • 2608.09853 • Published 2 days ago • 6
Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory Paper • 2608.07169 • Published 5 days ago • 31
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 2 days ago • 189
YOLO-PEFT: Parameter-Efficient Fine-Tuning on YOLO Family Paper • 2608.07051 • Published 5 days ago • 16
MASS: Multiplayer World Models with Authoritative Shared State Paper • 2608.06257 • Published 2 days ago • 14
From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models Paper • 2608.06020 • Published 6 days ago • 33
GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? Paper • 2608.05747 • Published 6 days ago • 43
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning Paper • 2608.05987 • Published 6 days ago • 90