ACLArena: Agent Continue Learning in Multi-stage Post-training Paper • 2609.23989 • Published 4 days ago • 11
onPanda: Efficient Annotation of On-Policy Alignment Data for LLMs and Agents via Token-Level Correction Paper • 2609.24983 • Published 4 days ago • 52
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 4 days ago • 196
Why Do Video Diffusion Models Violate Physics? Unveiling the Flaws in Attention Mechanisms Paper • 2609.23658 • Published 5 days ago • 23
One to More, More to One: Category-Aware Iterative Expert Training for Software Engineering Agents Paper • 2609.23377 • Published 5 days ago • 45
HuRo: Robotizing Human Videos for Scalable VLA Pretraining Paper • 2609.10706 • Published 7 days ago • 27
WorldCrafter: Consistent Video World Model with Implicit 3D-aware Memory Paper • 2609.24984 • Published 4 days ago • 145
Paint-Anything: Unified Any-Color Control for Image Generation and Editing Paper • 2609.20816 • Published 8 days ago • 54
GraphSkillEvo: Evolutionary Optimization of Graph-Structured Agent Skills Paper • 2609.21749 • Published 7 days ago • 17
HypoEvolve: Genetic Algorithms Enable Multi-Agent LLMs to Discover Scientific Hypotheses Paper • 2609.15938 • Published 11 days ago • 31