VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning Paper • 2608.26105 • Published Aug 26 • 193
PhyGround: Benchmarking Physical Reasoning in Generative World Models Paper • 2605.10806 • Published May 11 • 4
PAVXploreRL: Physical-Action-Visual World Model Reinforcement Learning with Action Exploration Paper • 2607.16602 • Published 26 days ago • 1
WorldExam: Benchmarking World Models from Apparent Appearance to Inherent Reactivity Paper • 2608.02603 • Published Aug 3 • 35
WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models Paper • 2608.04964 • Published Aug 5 • 15
AlayaWorld: Interactive Long-Horizon World Modeling -- Full Technical Report Paper • 2607.18367 • Published Jul 20 • 63
4DStreamCtrl: Interactive Video Generation with Online 4D Control Paper • 2608.25479 • Published Aug 27 • 2
HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis Paper • 2607.17097 • Published Jul 19 • 13
The Other Half of the Memory Wall: Serving 35B MoEs from SSD with Trained Routing Prediction Paper • 2609.18063 • Published 20 days ago • 22
DynaPix: Can Vision-Language Models Identify the Exact Future? Paper • 2608.05505 • Published Aug 6 • 1
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published Sep 3 • 28
OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping Paper • 2607.00881 • Published Jul 1 • 1
HumanMoveVQA: Can Video MLLMs reason about human movement in videos? Paper • 2606.27999 • Published Jun 29 • 1
GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? Paper • 2608.05747 • Published Aug 6 • 48
VGIF-Score: Interpretable and Diagnostic Evaluation of Spatio-Temporal Instruction Following in Video Generation Paper • 2607.13527 • Published Jul 15 • 2
VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing Paper • 2608.05485 • Published Aug 6 • 4