Gaze as Evidence for Common Grounding: A Cross-Corpus Analysis of MapTask and MUNDEX Paper • 2609.18011 • Published 2 days ago • 22
X-AuT: Progressive Audio-Encoder Compression for Speech LLMs with Cross-Scale Distillation Paper • 2609.11412 • Published 8 days ago • 46
An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics Paper • 2609.10712 • Published 9 days ago • 43
SceneMosaic: Efficient and Diverse Simulation-Ready Scene Generation via Hybrid Agentic Layout Evolution Paper • 2609.05594 • Published 14 days ago • 47
orcarouter/Qwen3.8-Flash-Next-Uncensored-GGUF Image-Text-to-Text • 177B • Updated 7 days ago • 198k • 345
4DAnyone: Create Anyone in 4D from a Casual Monocular Video Paper • 2608.20335 • Published 29 days ago • 84
Specification-first convergence with an AI coding agent: a case study of dismantling a core architectural invariant across 189 files in a 717k-line codebase with no test oracle and no human code review Paper • 2608.12440 • Published Aug 12 • 11
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 264
Ex-Omni-2D: Expressive Omni-Modal Dialogue Models with Native Visual Presence Paper • 2608.10720 • Published Aug 11 • 16
DyPES-VLA: Learning Shared Dynamics Priors and Embodiment-Specific Control for Cross-Embodiment Manipulation Paper • 2608.06374 • Published Aug 6 • 23
Ego2Robot: Scalable Robot Data Synthesis from Egocentric Human Data Paper • 2608.02580 • Published Aug 3 • 27
JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion Paper • 2608.03974 • Published Aug 4 • 106