MaLiang-Harness: A Programmable Path to Image and Video Generation Paper • 2609.34309 • Published 4 days ago • 384
VisionHOPE: Visual Backbones as Self-Modifying Learning Systems Paper • 2609.33325 • Published 5 days ago • 304
Beyond Dyadic Memory: Interaction-Aware Multimodal Memory with Adaptive Agentic Retrieval for Multi-Party Spoken Conversations Paper • 2609.32522 • Published 6 days ago • 70
YuE2: Unifying Symbolic and Audio Music Generation at Frontier Quality Paper • 2609.33757 • Published 5 days ago • 233
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published 22 days ago • 705
NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction Paper • 2609.10715 • Published 23 days ago • 331