HuRo: Robotizing Human Videos for Scalable VLA Pretraining Paper • 2609.10706 • Published 4 days ago • 8
Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models Paper • 2503.08737 • Published Mar 11, 2025
Scenes as Objects, Not Primitives: Instance-Structured 3D Tokenization from Unposed Views Paper • 2606.29513 • Published Jun 28 • 52
Scenes as Objects, Not Primitives: Instance-Structured 3D Tokenization from Unposed Views Paper • 2606.29513 • Published Jun 28 • 52
Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs Paper • 2507.07990 • Published Jul 10, 2025 • 45