UniSwap: Streaming Audio-Visual Identity Swapping for Talking Videos Paper • 2608.11752 • Published 5 days ago • 21
LiveAnimate: Stable Long-Form Streaming Human Animation in Real-Time Paper • 2608.11745 • Published 5 days ago • 22
AVA-Encoder: Towards Agent-Native Video Representation Learning Paper • 2608.12313 • Published 6 days ago • 14
What to Edit Next: Visually Aligned Image-Editing Follow-Up Suggestions in Conversational Systems Paper • 2608.07565 • Published 15 days ago • 28
Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation Paper • 2607.24731 • Published 22 days ago • 77
Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation Paper • 2607.05382 • Published Jul 9 • 88
CollectionLoRA: Collecting 50 Effects in 1 LoRA via Multi-Teacher On-Policy Distillation Paper • 2605.25378 • Published May 25 • 62
CollectionLoRA: Collecting 50 Effects in 1 LoRA via Multi-Teacher On-Policy Distillation Paper • 2605.25378 • Published May 25 • 62
RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time Paper • 2604.11626 • Published Apr 13 • 103
Stable-Makeup: When Real-World Makeup Transfer Meets Diffusion Model Paper • 2403.07764 • Published Mar 12, 2024 • 1
Stable-Hair v2: Real-World Hair Transfer via Multiple-View Diffusion Model Paper • 2507.07591 • Published Jul 10, 2025
Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Paper • 2512.04677 • Published Dec 4, 2025 • 179
PhotoDoodle: Learning Artistic Image Editing from Few-Shot Pairwise Data Paper • 2502.14397 • Published Feb 20, 2025 • 41
EasyControl: Adding Efficient and Flexible Control for Diffusion Transformer Paper • 2503.07027 • Published Mar 10, 2025 • 30
FonTS: Text Rendering with Typography and Style Controls Paper • 2412.00136 • Published Nov 28, 2024 • 1