Rethinking Long-Video Efficiency: A Joint Allocation Perspective on Frames, Pixels, and Front-End Latency Paper • 2610.04318 • Published 6 days ago • 18
Where Does Retrieval-Based Open-Ended Evaluation Fail? Automatic Taxonomy Induction from Long-Form Medical Answer Factuality Verification Paper • 2609.30467 • Published 15 days ago • 22
Learning What to Recall: Adaptive Multi-Cue Episodic Memory for World Models Paper • 2609.34677 • Published 11 days ago • 13
Multi-Grid Post-Training for Long-Form Multi-Shot Video Generation Paper • 2609.06373 • Published Sep 6 • 17
Using Grounded Theory for Agent Behavior Analysis at Scale Paper • 2608.30391 • Published Aug 31 • 19
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published Sep 3 • 186
Inject, Align, Recover: Staged Post-Training for Retrieval-Free Document Knowledge Internalization Paper • 2608.20281 • Published Aug 20 • 15