All modalities are equal, but video is more equal: Closing the Cross-Attention Gap in Joint Video Generation Paper • 2609.27901 • Published 6 days ago • 15
HARMONY: Hierarchical Agentic Reasoning for MONocular Image-to-Scene Synthesis Paper • 2609.26793 • Published 7 days ago • 7
Harness-Zero: Harness Distillation via Agent-as-Harness Paper • 2609.24974 • Published 8 days ago • 37
onPanda: Efficient Annotation of On-Policy Alignment Data for LLMs and Agents via Token-Level Correction Paper • 2609.24983 • Published 8 days ago • 55
CARE: Experience-Guided Atomic Corrective Execution for Vision-Language-Action Policies Paper • 2609.24118 • Published 8 days ago • 28
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 8 days ago • 211
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 12 days ago • 188
PANORAMA: Panoptic Grounded Captioning via Mask Proposal Selection Paper • 2609.19143 • Published 13 days ago • 18