RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 3 days ago • 177
CodeMidas: Scaling Agentic Coding RL Environments from Code Itself Paper • 2609.22068 • Published 6 days ago • 129
MindZero: Learning Online Mental Reasoning With Zero Annotations Paper • 2606.00240 • Published May 29 • 4
ThoughtTrace: Understanding User Thoughts in Real-World LLM Interactions Paper • 2605.20087 • Published May 19 • 17
AutoToM: Automated Bayesian Inverse Planning and Model Discovery for Open-ended Theory of Mind Paper • 2502.15676 • Published Feb 21, 2025 • 3