view article Article FLUX 3 Model Overview: Multimodal Flow Models for Image, Video, Audio, and Action Prediction ResterChed • Jul 24 • 25
view article Article 🧨 Accelerating Stable Diffusion XL Inference with JAX on Cloud TPU v5e +6 pcuenq, jffacevedo, alexspiridonov, pmotter, yyetim, svaibhav, vjsingh, patrickvonplaten • Oct 3, 2023 • 10
view article Article Training a coding model to paint watercolours with TRL and OpenEnv sergiopaniego • 19 days ago • 71
view article Article Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident +2 hlarcher, XciD, raphael-gl, chris-rannou • Jul 27 • 504
view article Article LightOnOCR-2-1B: a lightweight high-performance end-to-end OCR model family lightonai • Jan 19 • 104
view article Article How we OCR'ed 30,000 papers using Codex, open OCR models and Jobs nielsr • Apr 7 • 62
view article Article ⚡ nano-vLLM: Lightweight, Low-Latency LLM Inference from Scratch zamal • Jun 28, 2025 • 47
view article Article Assisted Generation: a new direction toward low-latency text generation joaogante • May 11, 2023 • 81
view article Article A Gentle Introduction to 8-bit Matrix Multiplication for transformers at scale using transformers, accelerate and bitsandbytes ybelkada, timdettmers • Aug 17, 2022 • 140
view article Article Continuous batching from first principles +1 ror, ArthurZ, mcpotato • Nov 25, 2025 • 444
SmolVLM: Redefining small and efficient multimodal models Paper • 2504.05299 • Published Apr 7, 2025 • 212
Neural Vocoder is All You Need for Speech Super-resolution Paper • 2203.14941 • Published Mar 28, 2022 • 1
MusicInfuser: Making Video Diffusion Listen and Dance Paper • 2503.14505 • Published Mar 18, 2025 • 12
view article Article Open-Source Handwritten Signature Detection Model samuellimabraz • Mar 14, 2025 • 124
view article Article Welcome Gemma 3: Google's all new multimodal, multilingual, long context open LLM +2 ariG23498, merve, pcuenq, reach-vb • Mar 12, 2025 • 499
Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities Paper • 2503.03983 • Published Mar 6, 2025 • 30
view article Article Using LoRA for Efficient Stable Diffusion Fine-Tuning pcuenq, sayakpaul • Jan 26, 2023 • 84