WanPE: Towards Cinematic Prompt Enhancement for Modern Text-to-Video Generation Paper • 2609.30221 • Published 6 days ago • 46
Geometric and Semantic Coupling for Interaction Understanding in 3D Scenes Paper • 2609.25247 • Published 9 days ago • 10
Can MiniMax-H3 Reason About the Physical World? An Evaluation of Omni-Modal Generative Model Paper • 2609.18323 • Published 14 days ago • 132
JEV-as-a-Judge: Accept When Confident, Escalate When Unsure Paper • 2609.26550 • Published 8 days ago • 42
Tri-PvP: Exposing Modality Bias in Omni-Modal Large Language Models through Perceptual-Propositional Evidence Conflicts Paper • 2609.06011 • Published 25 days ago • 18
ALPINE: Adaptive Localization for Parameter- and Sample-Efficient Few-Shot Learning Paper • 2609.22323 • Published 14 days ago • 10
Think Like a World Model, Act Like a VLA: Distilling World-Model Representations into Compact Robot Policies Paper • 2609.24682 • Published 9 days ago • 12
RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents Paper • 2609.22000 • Published 12 days ago • 79
WeVisDoc: From Coverage to Capability for Robust End-to-End Document Parsing Paper • 2609.20423 • Published 13 days ago • 48
Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening Paper • 2609.18708 • Published 14 days ago • 82
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 19 days ago • 264
WearableQA: A Benchmark for Health Reasoning over Real-World Wearable Data Paper • 2609.05405 • Published 26 days ago • 45
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 23 days ago • 375