Progressive Agent Skill Generation via Reinforcement Learning Paper • 2608.01678 • Published 2 days ago • 49
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published 10 days ago • 89
Meshy T2: Fast Native Mesh Generation with Flow Matching Paper • 2607.28675 • Published 8 days ago • 47
QQWorld: Quantile-Quantile Matching for World Model Regularization Paper • 2607.28415 • Published 6 days ago • 27
Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering Paper • 2607.28568 • Published 6 days ago • 179
HiFi-UMI: Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone Paper • 2607.25895 • Published 8 days ago • 152
Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation Paper • 2607.27372 • Published 7 days ago • 17
StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation Paper • 2607.26754 • Published 7 days ago • 18
CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization Paper • 2607.25659 • Published 8 days ago • 82
Pass the Baton: Trajectory-Relayed On-Policy Distillation Paper • 2607.26057 • Published 8 days ago • 33
Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering Paper • 2607.21848 • Published 13 days ago • 9
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing Paper • 2607.19064 • Published 15 days ago • 76
Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning Paper • 2607.18722 • Published 15 days ago • 35
EvolvingWorld: An Open-Schema Framework for Co-Evolving Role-Play Agents and World Model in Interactive Literary World Paper • 2607.17250 • Published 17 days ago • 92
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU Paper • 2607.19191 • Published 14 days ago • 308
HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enchancement Paper • 2607.18217 • Published 16 days ago • 61
Environment-free Synthetic Data Generation for API-Calling Agents Paper • 2607.16900 • Published 18 days ago • 20
VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders Paper • 2607.14088 • Published 21 days ago • 14