Disentangling Representation Evolution in Transformers through Directional Decomposition
Paper • 2609.15975 • Published • 4
Efficient and adaptive foundation models across language and multimodal intelligence.
Drop-Then-Recovery: How Redundant Are Vision-Language-Action Models?
Demystifying When Pruning Works via Representation Hierarchies