Human Cognition in Machines: A Unified Perspective of World Models Paper • 2604.16592 • Published Apr 17
Flash-WAM: Modality-Aware Distillation for World Action Models Paper • 2606.05254 • Published Jun 3 • 7
PhyGround: Benchmarking Physical Reasoning in Generative World Models Paper • 2605.10806 • Published May 11 • 3
ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models Paper • 2605.24011 • Published May 19 • 2
ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models Paper • 2605.24011 • Published May 19 • 2
ThinkJEPA: Empowering Latent World Models with Large Vision-Language Reasoning Model Paper • 2603.22281 • Published Mar 23 • 20
Ref-Adv: Exploring MLLM Visual Reasoning in Referring Expression Tasks Paper • 2602.23898 • Published Feb 27 • 10
Fine-T2I: An Open, Large-Scale, and Diverse Dataset for High-Quality T2I Fine-Tuning Paper • 2602.09439 • Published Feb 10 • 14
VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting Paper • 2507.05116 • Published Jul 7, 2025
CircuitSense: A Hierarchical Circuit System Benchmark Bridging Visual Comprehension and Symbolic Reasoning in Engineering Design Process Paper • 2509.22339 • Published Sep 26, 2025 • 1
AnalogGenie: A Generative Engine for Automatic Discovery of Analog Circuit Topologies Paper • 2503.00205 • Published Feb 28, 2025
PhD Knowledge Not Required: A Reasoning Challenge for Large Language Models Paper • 2502.01584 • Published Feb 3, 2025 • 9
Themis: Towards Flexible and Interpretable NLG Evaluation Paper • 2406.18365 • Published Jun 26, 2024
NNsight and NDIF: Democratizing Access to Foundation Model Internals Paper • 2407.14561 • Published Jul 18, 2024 • 35