Ego2Act: Evaluating Goal-Directed Manipulation in Egocentric Video Generation Paper • 2610.01092 • Published 6 days ago • 33
AutoGUIWorld: Image Generators as Visual World Models for GUI Agent Paper • 2610.01215 • Published 6 days ago • 61
Adaptive Reward Routing: Dynamic Multi-Reward Optimization for Joint Audio-Video Diffusion via Forward-Process RL Paper • 2609.37200 • Published 8 days ago • 138
Beyond Memory: Harnessing Long-Horizon Agents with Explicit Belief States Paper • 2610.01415 • Published 6 days ago • 94
World Observer: Joint Actor-Observer Generation for Persistent World Modeling Paper • 2610.02162 • Published 6 days ago • 85
4Director: Controlling Video World Models with Rigid 3D Geometry Paper • 2610.02160 • Published 6 days ago • 38
Predictive Credit: Measuring What Scientific Explanations Add to Experimental Forecasts Paper • 2610.00314 • Published 8 days ago • 101
RoboCoach: World Models as Active Coaches for Compositional Robot Skills Paper • 2609.39685 • Published 7 days ago • 15
Physis-Lang: Self-Evolving Language as a Physical Representation for Video World Model Paper • 2609.40358 • Published 7 days ago • 19
ViRDM: Taming Representation Distribution Matching for Few-Step Causal Video Generation Paper • 2609.28923 • Published 13 days ago • 11
EVOKE: Eliciting World Knowledge in Agents for Transferable Decision-Making Paper • 2609.38334 • Published 6 days ago • 79
BiasReducer: Adaptive Bias Mitigation for Reward Models Paper • 2609.32720 • Published 11 days ago • 25
CompoWorld: Compositional Environment Scaling for General Agents Paper • 2609.33665 • Published 10 days ago • 42
Agent-Editing World Model: Rethinking World Modeling for LLM Agents Paper • 2609.28416 • Published 14 days ago • 43
Precise Editing and Flexible Referencing for Interactable Worlds Paper • 2609.34470 • Published 9 days ago • 20
Think Before You Score: Thinking Reward Model for Visual Generation Paper • 2609.37372 • Published 8 days ago • 101
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 10 days ago • 567
Anisotropic Representations Improve Planning in JEPA World Models Paper • 2609.37441 • Published 8 days ago • 34
WorldAttention: An Efficient Attention Architecture for Interactive Video World Models Paper • 2609.34606 • Published 9 days ago • 49