General Multimodal Protein Design Enables DNA-Encoding of Chemistry Paper • 2604.05181 • Published Apr 6 • 31
XFACTORS: Disentangled Information Bottleneck via Contrastive Supervision Paper • 2601.21688 • Published Jan 29
Revealing Subtle Phenotypes in Small Microscopy Datasets Using Latent Diffusion Models Paper • 2502.09665 • Published Feb 12, 2025
CubeComposer: Spatio-Temporal Autoregressive 4K 360° Video Generation from Perspective Video Paper • 2603.04291 • Published Mar 4 • 16
CubeComposer: Spatio-Temporal Autoregressive 4K 360° Video Generation from Perspective Video Paper • 2603.04291 • Published Mar 4 • 16
SemanticMoments: Training-Free Motion Similarity via Third Moment Features Paper • 2602.09146 • Published Feb 9 • 21
BFTBrain: Adaptive BFT Consensus with Reinforcement Learning Paper • 2408.06432 • Published Aug 12, 2024
DLFR-VAE: Dynamic Latent Frame Rate VAE for Video Generation Paper • 2502.11897 • Published Feb 17, 2025
SALAD: Achieve High-Sparsity Attention via Efficient Linear Attention Tuning for Video Diffusion Transformer Paper • 2601.16515 • Published Jan 23 • 15
SALAD: Achieve High-Sparsity Attention via Efficient Linear Attention Tuning for Video Diffusion Transformer Paper • 2601.16515 • Published Jan 23 • 15
CaricatureGS: Exaggerating 3D Gaussian Splatting Faces With Gaussian Curvature Paper • 2601.03319 • Published Jan 6 • 54
DiffusionBrowser: Interactive Diffusion Previews via Multi-Branch Decoders Paper • 2512.13690 • Published Dec 15, 2025 • 3
Efficiently Reconstructing Dynamic Scenes One D4RT at a Time Paper • 2512.08924 • Published Dec 9, 2025 • 25
Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer Paper • 2511.22699 • Published Nov 27, 2025 • 249
Decoupled DMD: CFG Augmentation as the Spear, Distribution Matching as the Shield Paper • 2511.22677 • Published Nov 27, 2025 • 36
Time-to-Move: Training-Free Motion Controlled Video Generation via Dual-Clock Denoising Paper • 2511.08633 • Published Nov 9, 2025 • 58
view post Post 33993 Want to iterate on a Hugging Face Space with an LLM? Now you can easily convert any HF entire repo (Model, Dataset or Space) to a text file and feed it to a language model! multimodalart/repo2txt See translation 3 replies · 🤗 3 3 👍 3 3 🚀 2 2 🧠 1 1 + Reply