Latent-Foresight: End-to-End Learning Predictable Representations for Latent World Models Paper • 2610.01942 • Published 5 days ago • 8
RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning Paper • 2609.20784 • Published 19 days ago • 57
ProgramDistill: From Interactive Web Apps to Verifiable Reference-Guided SWE Tasks Paper • 2609.18805 • Published 20 days ago • 67
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 29 days ago • 376
Kalman Delta Networks: Uncertainty-aware Associative Memory Paper • 2609.07816 • Published 29 days ago • 29