For-Value: Efficient Forward-Only Data Valuation for finetuning LLMs and VLMs
Paper • 2508.10180 • Published • 19
None defined yet.
Learning to Trigger: Reinforcement Learning at the Large Hadron Collider
A Theoretical Framework for Auxiliary-Loss-Free Load Balancing of Sparse Mixture-of-Experts in Large-Scale AI Models