LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling Paper • 2605.08083 • Published 4 days ago • 57
Mean Mode Screaming: Mean--Variance Split Residuals for 1000-Layer Diffusion Transformers Paper • 2605.06169 • Published 5 days ago • 111
Mela: Test-Time Memory Consolidation based on Transformation Hypothesis Paper • 2605.10537 • Published 1 day ago • 6