AlayaWorld: Interactive Long-Horizon World Modeling -- Full Technical Report Paper • 2607.18367 • Published 2 days ago • 43
DiffGI: Differentiable Geometry Images for High-Fidelity Thin-Shell 3D Generation Paper • 2607.13365 • Published 8 days ago • 13
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 9 days ago • 209
4D Human-Scene Reconstruction from Low-Overlap Captures Paper • 2607.09125 • Published 13 days ago • 53
Video Generation Models are General-Purpose Vision Learners Paper • 2607.09024 • Published 13 days ago • 83
Video-Oasis: Rethinking Evaluation of Video Understanding Paper • 2603.29616 • Published 21 days ago • 64
Why Can't I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition Paper • 2601.16211 • Published 21 days ago • 54
SceneFrom3D: Geometry-Conditioned Outdoor 3D Scene Generation via View Scheduling with Object-Level Control Paper • 2607.04540 • Published 18 days ago • 4
LLM-as-a-Tutor: Policy-Aware Prompt Adaptation for Non-Verifiable RL Paper • 2607.04412 • Published 18 days ago • 34
3D HAMSTER: Bridging Planning and Control in Hierarchical Vision Language Action Models through 3D Trajectory Guidance Paper • 2606.31329 • Published 23 days ago • 7
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published 25 days ago • 168
Program-as-Weights: A Programming Paradigm for Fuzzy Functions Paper • 2607.02512 • Published 21 days ago • 122
Domain Arithmetic: One-Shot VLA Adaptation under Environmental Shifts Paper • 2607.00666 • Published 22 days ago • 24
Are We Measuring Strategy or Phrasing? The Gap Between Surface- and Approach-Level Diversity in LLM Math Reasoning Paper • 2606.29985 • Published 24 days ago • 20
Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning Paper • 2503.15558 • Published Mar 18, 2025 • 52