Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making Paper • 2607.14277 • Published 14 days ago • 8
A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever Paper • 2607.23806 • Published 3 days ago • 4
Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification Paper • 2607.24027 • Published 2 days ago • 28
IDEAgent: Agentic Quality-Diversity Search for Research Idea Generation Paper • 2607.22375 • Published 5 days ago • 8
Interactive Training 2: Auditable Control Plane for Live Model Training Paper • 2607.18314 • Published 12 days ago • 18
Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems Paper • 2607.21503 • Published 6 days ago • 22
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Paper • 2607.21653 • Published 7 days ago • 29
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Paper • 2607.22529 • Published 5 days ago • 37
Stepwise Reasoning Enhancement for LLMs via External Subgraph Generation Paper • 2606.04454 • Published Jun 3 • 1
DeLIVeR: Decomposed Learning for Information-grounded Veracity Recognition via Reinforced Knowledge Graph Exploration Paper • 2607.17935 • Published 9 days ago • 1
Efficient Retrieval-Augmented Generation via Token Co-occurrence Graphs Paper • 2606.30093 • Published about 1 month ago • 1
All Relations Lead to Rome: Automated Knowledge Graph Creation and Question Generation Paper • 2606.22645 • Published Jun 21 • 1
SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation Paper • 2607.21553 • Published 6 days ago • 37
AoiZora: Topology-Aware Auto-Parallel Optimization for Inference of Diffusion Transformers Paper • 2606.17566 • Published Jun 16 • 1
HyperVAttention: Efficient Sparse Attention with Spatio-Temporal Clustering for Video Diffusion Paper • 2607.03012 • Published 26 days ago • 1
ScalingAttention: Discovering Intrinsic Sparse Attention Topology for Video Diffusion Transformers Paper • 2606.23019 • Published Jun 22 • 1
Predict, Reuse, and Repair: Accelerating Dynamic Sparse Attention for Long-Context LLM Decoding Paper • 2606.30389 • Published about 1 month ago • 1
Chorus II: Cross-Request Sparsity Reuse for Efficient Image-to-Video Generation Paper • 2606.25040 • Published Jun 23 • 1
OSP-Next: Efficient High-Quality Video Generation with Sparse Sequence Parallelism, HiF8 Quantization, and Reinforcement Learning Paper • 2605.28691 • Published May 27 • 25