VepAgent: Bridging Causal-Transition via Tool-Augmented Reinforcement Learning for Video Event Prediction Paper • 2610.06293 • Published 4 days ago • 51
Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement Paper • 2609.13406 • Published 28 days ago • 85
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 25 days ago • 252
trl-internal-testing/tiny-Qwen2ForCausalLM-2.5 Text Generation • 2.44M • Updated about 1 month ago • 7.5M • 56
NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction Paper • 2609.10715 • Published 30 days ago • 331
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published Aug 12 • 111
SimWAM: A Simple World Action Model for End-to-End Autonomous Driving Paper • 2608.07468 • Published Aug 7 • 41
DistillAlign: Coordinating Mode Covering and Mode Seeking in Autoregressive Video Distillation Paper • 2607.26811 • Published Jul 29 • 18
Progress Reward Modeling for Robotic Learning: A Comprehensive Survey Paper • 2607.21655 • Published Jul 22 • 112