Change the Product, Keep the Parameters: Associative Algebra Layers for Transformers Paper • 2609.32814 • Published 4 days ago • 17
DeepMostInnovations/sales-conversion-model-reinf-learning Reinforcement Learning • Updated 13 days ago • 568 • 210
From Production Traffic to Post-Training: Building a Self-Hosted LLM That Covers the Corporate Request Mix Paper • 2609.01572 • Published 29 days ago • 36
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation Paper • 2511.14993 • Published Nov 19, 2025 • 236
Reinforcement Learning for Reasoning in Large Language Models with One Training Example Paper • 2504.20571 • Published Apr 29, 2025 • 99
Running 4.05k The Ultra-Scale Playbook 🌌 4.05k The ultimate guide to training LLM on large GPU Clusters
Running 602 Scaling test-time compute 📈 602 Boost LLM answers with flexible test‑time search strategies
AnatoliiPotapov/T-lite-instruct-0.1 Text Generation • 8B • Updated Sep 25, 2024 • 82 • • 103