view article Article Transformers now runs llama.cpp quants +1 marcsun13, ArthurZ, lysandre • 5 days ago • 71
LLM Compression by Block Removal with Constrained Binary Optimization Paper • 2602.00161 • Published Jun 17 • 9
view article Article Pruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem MultiverseComputingCAI • 6 days ago • 31
Compile by Training: Turning Natural-Language Specifications into Local Neural Functions Paper • 2609.04199 • Published 24 days ago • 328
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 24 days ago • 186
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 24 days ago • 138
view article Article LFM2.5 Q4\_0 Checkpoints from Quantization-Aware Distillation LiquidAI • Aug 19 • 36
Weak-to-Strong Generalization via Direct On-Policy Distillation Paper • 2607.05394 • Published Jul 8 • 144
view article Article Distillation in 2026 (so far): which frontier models use it and how sergiopaniego • Jul 8 • 22
view article Article Making Knowledge Distillation Cheap Enough to Run at Scale MultiverseComputingCAI • Aug 10 • 42
view article Article VLX-Seek: Improving VLM Fine-Grained Perception via Region Reference Instead of Coordinate Generation omlab • Jun 27 • 14
Can Current Agents Close the Discovery-to-Application Gap? A Case Study in Minecraft Paper • 2604.24697 • Published Apr 27 • 2
LLaVA-UHD v4: What Makes Efficient Visual Encoding in MLLMs? Paper • 2605.08985 • Published May 9 • 23