Avatar-Forever: Decoupled Parallel Training for High-Quality Real-Time Infinite Avatars Paper • 2608.12107 • Published Aug 12 • 3
Agent Lightning: Train ANY AI Agents with Reinforcement Learning Paper • 2508.03680 • Published Aug 5, 2025 • 142
Smaller Models are Natural Explorers for Policy-Level Diversity in GRPO Paper • 2605.30789 • Published Jun 2 • 27
Smaller Models are Natural Explorers for Policy-Level Diversity in GRPO Paper • 2605.30789 • Published Jun 2 • 27
AnchorWorld: Embodied Egocentric World Simulation with View-based Evolution Customization Paper • 2606.07326 • Published Jun 5 • 30
Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation Paper • 2605.28091 • Published May 27 • 6
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments Paper • 2605.30280 • Published May 28 • 146
Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation Paper • 2605.28091 • Published May 27 • 6
LongLive-2.0: An NVFP4 Parallel Infrastructure for Long Video Generation Paper • 2605.18739 • Published May 18 • 117
Beyond Length Scaling: Synergizing Breadth and Depth for Generative Reward Models Paper • 2603.01571 • Published Mar 2 • 34
DP$^2$O-SR: Direct Perceptual Preference Optimization for Real-World Image Super-Resolution Paper • 2510.18851 • Published Oct 21, 2025