EffectLearner: World-Aware Object-Effect Reasoning for Real-World Video Object Removal Paper • 2608.05565 • Published 1 day ago • 12
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper • 2607.27919 • Published 9 days ago • 58
Tooony133/dinov3-vits16-pretrain-lvd1689m Image Feature Extraction • 21.6M • Updated Jun 19 • 1.38k • 1
TimeLens2 Collection Generalist Video Temporal Grounding with Multimodal LLMs • 8 items • Updated 10 days ago • 15
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Paper • 2607.14935 • Published 23 days ago • 172
World-R1: Reinforcing 3D Constraints for Text-to-Video Generation Paper • 2604.24764 • Published Apr 27 • 119