SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published 13 days ago • 103
Mastermind: Strategy-grounded Learning for Repository-Scale Vulnerability Reproduction Paper • 2607.01764 • Published 27 days ago • 8
A Subgoal-driven Framework for Improving Long-Horizon LLM Agents Paper • 2603.19685 • Published Mar 20 • 22