WanPE: Towards Cinematic Prompt Enhancement for Modern Text-to-Video Generation Paper • 2609.30221 • Published 6 days ago • 46
SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 20 days ago • 275
Scaling Properties of Text Conditioning in Visual Generation Paper • 2607.29679 • Published Jul 31 • 42
Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation Paper • 2607.11886 • Published Jul 13 • 61
Bernini: Latent Semantic Planning for Video Diffusion Paper • 2605.22344 • Published May 21 • 21
CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives Paper • 2605.12496 • Published May 12 • 31
WRBench: Current World Models Lack a Persistent State Core Collection WRBench public release: paper, prompts, videos, scores, human labels, and leaderboard. • 6 items • Updated Jul 7 • 3
Memento: Reconstruct to Remember for Consistent Long Video Generation Paper • 2606.14667 • Published Jun 12 • 19
VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior Tracing Paper • 2605.30117 • Published May 28 • 2