Rethinking Cross-Tokenizer On-Policy Distillation: From Alignment Coverage to Supervision Reliability Paper • 2610.08448 • Published 2 days ago • 153
DuoMatching: Joint-Marginal Distribution Matching for Few-Step Video Generation Paper • 2610.03543 • Published 6 days ago • 68
EVISKILL: Grounding Skill Evolution in Replayable Evidence Paper • 2610.05030 • Published 4 days ago • 36
DistScene: Object-to-Scene Distillation for 3D Scene Generation Paper • 2610.06960 • Published 5 days ago • 5
VeriFine: Scaling Verification for Self-Improvement in Embodied Reasoning Paper • 2610.08761 • Published 2 days ago • 4
Selection-Based Structured Reasoning: Toward Efficient Multimodal Search Agents Paper • 2610.01892 • Published 7 days ago • 14
OSWorld-Pro: Process-based Evaluation for Computer Use Agents Paper • 2609.24890 • Published 16 days ago • 22
Dynamic Harness Search: Building Multi-Agent Systems Per-Query via Prediction Paper • 2610.04137 • Published 6 days ago • 10
Kandinsky 6.0 Video: Foundation Models for Synchronized Video and Audio Generation Paper • 2610.05608 • Published 4 days ago • 132
HelixWorld: A Real-time Interactive Audio-Visual World Model Paper • 2609.38123 • Published 9 days ago • 34
Skill2Real: Agentic Skill Learning for Zero-Shot Sim-to-Real Robot Manipulation Paper • 2610.02788 • Published 6 days ago • 17
Scaling Trajectories for Complex Tasks through Recursive Self-Rewrite Paper • 2610.02826 • Published 6 days ago • 99
EditHero: A Benchmark for Long-Horizon Part-Level 3D Editing and Vibe Modeling Paper • 2610.02298 • Published 7 days ago • 55
Transformers Stop Thinking Too Early, and a Tiny LoRA Fixes It Paper • 2609.36585 • Published 9 days ago • 76
World Observer: Joint Actor-Observer Generation for Persistent World Modeling Paper • 2610.02162 • Published 7 days ago • 85
Beyond Memory: Harnessing Long-Horizon Agents with Explicit Belief States Paper • 2610.01415 • Published 7 days ago • 94
ROWBench: Do Video Models Render What the Program Specifies? Paper • 2610.02205 • Published 7 days ago • 71