RunningTensor: Generalizing Linear Attention to Higher-Order Recurrent States Paper • 2609.12814 • Published 16 days ago • 1
Mahalanobis-Based Multi-Head Attention for Complex State Propagation Paper • 2608.24462 • Published Aug 25 • 1
Complex KDA: Understanding and Enhancing the Expressivity of Kimi Delta Attention Paper • 2609.24797 • Published 6 days ago • 11
Cross-Model Memory Transfer via Target-Side Reader Adaptation Paper • 2608.17050 • Published Aug 17 • 7
MemoryAthena: Adaptive Routing over Latent and Generated Memories Paper • 2609.25853 • Published 5 days ago • 5
PhysBrain 1.5: From Vision-Language Models to Physical Foundation Models Paper • 2609.14973 • Published 13 days ago • 174
FutureWorld: A Live Reinforcement Learning Environment for Predictive Agents with Real-World Outcome Rewards Paper • 2604.26733 • Published May 15 • 1
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 16 days ago • 263
Texture Generation on 3D Meshes with Point-UV Diffusion Paper • 2308.10490 • Published Aug 21, 2023 • 1
AA-SVD : Anchored and Adaptive SVD for Large Language Model Compression Paper • 2604.02119 • Published Apr 2 • 1
OASIS: Online Activation Subspace Learning for Memory-Efficient Training Paper • 2604.09406 • Published Apr 10 • 1
Running on Zero MCP 9 Qwen-Image-2.1-LoRAs-PnP 🚀 9 Collection of Qwen-Image-2.1 LoRAs — Plug and Play