arxiv:2605.11182
Xuyan Ye
LulaCola
AI & ML interests
LLM Reasoning, Self-Evolving Agent
Recent Activity
upvoted a paper about 2 hours ago
AgentDebugX: An Open-Source Toolkit for Failure Observability, Attribution, and Recovery in LLM Agents authored a paper 2 months ago
The Many Faces of On-Policy Distillation: Pitfalls, Mechanisms, and Fixes upvoted a paper 2 months ago
The Many Faces of On-Policy Distillation: Pitfalls, Mechanisms, and Fixes