When to Switch: Reliable Action-Chunk Extension for Vision-Language-Action Models Paper • 2610.05719 • Published 5 days ago • 31
Kandinsky 6.0 Video: Foundation Models for Synchronized Video and Audio Generation Paper • 2610.05608 • Published 6 days ago • 160
robot-learning-group47/eval_1_full_prompt_smolvla_blue_v2 SO-101 • Updated May 14 • 0 episodes • 459 • 6
MotorMind: Scaffolding General Vision Language Models for Zero-Shot Robot Manipulation Paper • 2609.38078 • Published 11 days ago • 111
Does Learning Protein Folding Generalize to Broader Reasoning? Paper • 2609.38879 • Published 10 days ago • 61
VoxMem: Benchmarking Multimodal Memory in Large Audio Language Models Paper • 2609.32607 • Published 14 days ago • 134
EVO-WAM: Evolving World Action Models through Video-Action Verification Paper • 2609.38057 • Published 11 days ago • 49
WorldLine: Action-Driven Visual Simulation for Robotic Manipulation Paper • 2609.38059 • Published 11 days ago • 33
In-Context Learning for Robots: Methods and Applications Paper • 2609.36012 • Published 12 days ago • 322