Change the Product, Keep the Parameters: Associative Algebra Layers for Transformers Paper • 2609.32814 • Published 4 days ago • 17
Risk-Averse Reinforcement Learning with Itakura-Saito Loss Paper • 2505.16925 • Published May 22, 2025 • 26
Linear Transformers with Learnable Kernel Functions are Better In-Context Models Paper • 2402.10644 • Published Feb 16, 2024 • 81