RAVEN: Real-time Autoregressive Video Extrapolation with Consistency-model GRPO Paper • 2605.15190 • Published May 14 • 14
view article Article State of Open Models: Summer 2026 Observations +1 AdinaY, multimodalart, irenesolaiman • 8 days ago • 138
SpotSound: Enhancing Large Audio-Language Models with Fine-Grained Temporal Grounding Paper • 2604.13023 • Published Apr 14 • 2
LTX-2.5 Collection LTX-2.5 base models, quantized models and accompanying LoRAs and IC-LoRAs • 4 items • Updated 10 days ago • 46
Reference-Driven Multi-Speaker Audio Scene Generation from In-the-Wild Priors Paper • 2606.19325 • Published Jun 17 • 2
Scaling Properties of Text Conditioning in Visual Generation Paper • 2607.29679 • Published 22 days ago • 40
view article Article IDEOGRAM-4 for inpainting with Modular Diffusers and Differential Diffusion OzzyGT • 16 days ago • 6
QueenVIS: Rethinking Image-Only Training for Video Instance Segmentation via Query Enrichment Paper • 2607.24598 • Published 26 days ago • 9
StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation Paper • 2607.26754 • Published 24 days ago • 18
view article Article Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident +2 hlarcher, XciD, raphael-gl, chris-rannou • 26 days ago • 478
Laguna S 2.1 Collection Our most capable model to date, designed for long-horizon work. • 13 items • Updated 18 days ago • 45
view article Article Experimenting with the proposed Cross-Origin Storage API in Transformers.js tomayac • Jun 23 • 8
view article Article Be Ready Before the Attack: A Practical Guide to Self-Hosting an Open Model for Cyber Defense jeffboudier • Jul 20 • 20