Distill Where You Fail: Recovering Learning Signals of Negative RL-Groups from Adaptive Teacher Guidance Paper • 2608.00782 • Published 6 days ago • 14
FocusMem: Factorizing Content, Readout, and Trust in Latent GUI Memory Paper • 2608.04530 • Published 2 days ago • 10
HelloWorld: Enabling Socially Interactive Characters in Video World Models Paper • 2608.05070 • Published 2 days ago • 17
WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models Paper • 2608.04964 • Published 2 days ago • 10
OPD-V: Visual On-Policy Self-Distillation with Modality Balance Paper • 2608.05131 • Published 2 days ago • 8
BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Paper • 2608.05042 • Published 2 days ago • 7
JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion Paper • 2608.03974 • Published 3 days ago • 83
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM Paper • 2607.27205 • Published 9 days ago • 139
Self Gradient Forcing: Native Long Video Extrapolation Paper • 2607.20368 • Published 16 days ago • 35
DreamForge-World 0.1 Preview: A Low-Compute Real-Time Controllable World Model Paper • 2606.30292 • Published Jun 29 • 16
JoyAI-VL-Interaction: Real-Time Vision-Language Interaction Intelligence Paper • 2606.14777 • Published Jun 10 • 216
Youtu-LLM: Unlocking the Native Agentic Potential for Lightweight Large Language Models Paper • 2512.24618 • Published Dec 31, 2025 • 155
TurboDiffusion: Accelerating Video Diffusion Models by 100-200 Times Paper • 2512.16093 • Published Dec 18, 2025 • 96
DreaMontage: Arbitrary Frame-Guided One-Shot Video Generation Paper • 2512.21252 • Published Dec 24, 2025 • 35
Details Collection A gated collection of datasets containing evaluation details • 4500 items • Updated Mar 2 • 6