Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization Paper • 2608.16072 • Published 6 days ago • 148 • 5
Optimizing Visual Generative Models via Distribution-wise Rewards Paper • 2607.02291 • Published Jul 2 • 18 • 3
IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation Paper • 2606.24849 • Published Jun 23 • 19 • 2
LatentSkill: From In-Context Textual Skills to In-Weight Latent Skills for LLM Agents Paper • 2606.06087 • Published Jun 4 • 67 • 8
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills Paper • 2606.07412 • Published Jun 5 • 12 • 3
LatentSkill: From In-Context Textual Skills to In-Weight Latent Skills for LLM Agents Paper • 2606.06087 • Published Jun 4 • 67 • 8
Think in Strokes, Not Pixels: Process-Driven Image Generation via Interleaved Reasoning Paper • 2604.04746 • Published Apr 8 • 73 • 4