On-Policy Parameter Update Direction Underlies Generalization in LLM Post-Training Paper • 2609.36659 • Published 9 days ago • 84
Srijika: OpenType-Layout-Reusing Font Restyling for Nine Indic Scripts Paper • 2609.05661 • Published Sep 4 • 41
Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening Paper • 2609.18708 • Published 22 days ago • 84
Why Gated DeltaNet Survives 4-Bit Quantization: NVFP4 W4A4 for the Recurrent Half of a Hybrid 27B LLM Paper • 2609.04098 • Published Sep 3 • 86
Are You Sure You're Sure? On the Impact of Instruction Tuning on Confidence and Lexical Diversity Paper • 2608.13430 • Published Aug 13 • 13
Task-Conditional Flow Matching for Balanced Multilingual Text Embedding Adaptation Paper • 2608.05785 • Published Aug 6 • 11
DyPES-VLA: Learning Shared Dynamics Priors and Embodiment-Specific Control for Cross-Embodiment Manipulation Paper • 2608.06374 • Published Aug 6 • 23
BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation Paper • 2608.05042 • Published Aug 5 • 9