UniEvo-VL: An On-policy Self-Distillation Training Recipe for Multimodal Model Self-improvement Paper • 2609.38721 • Published 9 days ago • 294
Transferring the Intelligence of VLMs to Robotic Control Paper • 2609.22966 • Published 20 days ago • 120
NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published Sep 8 • 328