DylanSh/so101_red_round_cube_blue_cup_currentcal_clean_train24_v1_20260811 Viewer • Updated 1 day ago • 19.8k • 66 • 1
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs Paper • 2608.01755 • Published 11 days ago • 140
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published 15 days ago • 302
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM Paper • 2607.27205 • Published 16 days ago • 139
Search and Refine During Think: Autonomous Retrieval-Augmented Reasoning of LLMs Paper • 2505.11277 • Published May 16, 2025 • 64
MuScriptor: An Open Model for Multi-Instrument Music Transcription Paper • 2607.08168 • Published Jul 9 • 21
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models Paper • 2606.11324 • Published Jun 9 • 172
InterleaveThinker: Reinforcing Agentic Interleaved Generation Paper • 2606.13679 • Published Jun 11 • 84
Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets Paper • 2605.28510 • Published May 27 • 5
Crafter: A Multi-Agent Harness for Editable Scientific Figure Generation from Diverse Inputs Paper • 2605.30611 • Published May 28 • 253