Self-Retrospection Distillation: Turning Post-hoc Experiences into Prior Foresight Paper • 2610.08077 • Published 4 days ago • 131
SGF+: Decoupling Gradient Flows for Autoregressive Video Generation Paper • 2610.10429 • Published 3 days ago • 55
Gains and Collapse in On-Policy Distillation:A Reinforcement Learning Perspective Paper • 2610.03185 • Published 8 days ago • 29
Questioning the Questions: Sustaining Self-Evolution in Reasoning Models Paper • 2610.04299 • Published 7 days ago • 68
Post-Training Frontier Text-to-Image Models by Composing Preference and Rubric Rewards Paper • 2610.02967 • Published 8 days ago • 23
Multi-Agent Egocentric World Model with Fine-Grained Embodied Interaction Paper • 2610.12299 • Published 2 days ago • 41
RSIGame: Autonomous Agentic Game Development with Recursive Self-improvement Paper • 2609.39045 • Published 10 days ago • 93
LimiX-2: A Contextual Mechanism Network Towards General Structured-Data Intelligence Paper • 2609.17488 • Published 25 days ago • 559
Mid-Harness: Scaling Actions Between Model and Harness for Terminal Agents Paper • 2609.39982 • Published 10 days ago • 119
The Teacher Is a Direction, Not a Destination: Extrapolating RL-Induced Representation Residuals in On-Policy Distillation Paper • 2609.36484 • Published 11 days ago • 397
Retrieval-Augmented Skill Optimization via Cross-Harness Adaptation Paper • 2609.38024 • Published 11 days ago • 64
World Observer: Joint Actor-Observer Generation for Persistent World Modeling Paper • 2610.02162 • Published 9 days ago • 88