Partition the Support, Reconstruct the Residual: Training-Free Sparse Attention for Video Generation and World Models Paper • 2608.18484 • Published 7 days ago • 8
Training a Student Expert via Semi-Supervised Foundation Model Distillation Paper • 2604.03841 • Published Apr 4 • 11