LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation Paper • 2608.30935 • Published 2 days ago • 25
LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation Paper • 2608.30935 • Published 2 days ago • 25
Deceptive-Human: Prompt-to-NeRF 3D Human Generation with 3D-Consistent Synthetic Images Paper • 2311.16499 • Published Nov 27, 2023 • 1
Agentic 3D Scene Generation with Spatially Contextualized VLMs Paper • 2505.20129 • Published May 26, 2025 • 1
Martian World Models: Controllable Video Synthesis with Physically Accurate 3D Reconstructions Paper • 2507.07978 • Published Jul 10, 2025
SmartAvatar: Text- and Image-Guided Human Avatar Generation with VLM AI Agents Paper • 2506.04606 • Published Jun 5, 2025
Trace Anything: Representing Any Video in 4D via Trajectory Fields Paper • 2510.13802 • Published Oct 15, 2025 • 31
CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving Paper • 2510.07944 • Published Oct 9, 2025 • 25
Trace Anything: Representing Any Video in 4D via Trajectory Fields Paper • 2510.13802 • Published Oct 15, 2025 • 31
PRELUDE: A Benchmark Designed to Require Global Comprehension and Reasoning over Long Contexts Paper • 2508.09848 • Published Aug 13, 2025 • 71
The Stochastic Parrot on LLM's Shoulder: A Summative Assessment of Physical Concept Understanding Paper • 2502.08946 • Published Feb 13, 2025 • 193