Memento: Reconstruct to Remember for Consistent Long Video Generation Paper • 2606.14667 • Published Jun 12 • 18
Memento: Reconstruct to Remember for Consistent Long Video Generation Paper • 2606.14667 • Published Jun 12 • 18
Memento: Reconstruct to Remember for Consistent Long Video Generation Paper • 2606.14667 • Published Jun 12 • 18
Sparse Growing Transformer: Training-Time Sparse Depth Allocation via Progressive Attention Looping Paper • 2603.23998 • Published Apr 16 • 1
Learning to Generate via Understanding: Understanding-Driven Intrinsic Rewarding for Unified Multimodal Models Paper • 2603.06043 • Published Mar 6
Blink: Dynamic Visual Token Resolution for Enhanced Multimodal Understanding Paper • 2512.10548 • Published May 23
V-ITI: Mitigating Hallucinations in Multimodal Large Language Models via Visual Inference-Time Intervention Paper • 2512.03542 • Published Dec 3, 2025
Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts Paper • 2509.21892 • Published May 11
Clip-Tuning: Towards Derivative-free Prompt Learning with a Mixture of Rewards Paper • 2210.12050 • Published Oct 21, 2022