moonshotai/Kimi-Linear-48B-A3B-Instruct Text Generation β’ 49B β’ Updated Dec 16, 2025 β’ 184k β’ β’ 584
view article Article Aligning to What? Rethinking Agent Generalization in MiniMax M2 MiniMax-AI β’ Oct 30, 2025 β’ 43
view article Article Why Did MiniMax M2 End Up as a Full Attention Model? MiniMax-AI β’ Oct 30, 2025 β’ 82
MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe Paper β’ 2509.18154 β’ Published Sep 16, 2025 β’ 63
meituan-longcat/LongCat-Flash-Chat Text Generation β’ 562B β’ Updated Sep 24, 2025 β’ 41.3k β’ 537
view article Article SmolLM3: smol, multilingual, long-context reasoner +21 eliebak, cmpatino, anton-l, edbeeching, m-ric, nouamanetazi, akseljoonas, guipenedo, hynky, clefourrier, SaylorTwift, kashif, qgallouedec, hlarcher, glutamatt, Xenova, reach-vb, ngxson, craffel, lewtun, loubnabnl, lvwerra, thomwolf β’ Jul 8, 2025 β’ 790
MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention Paper β’ 2506.13585 β’ Published Jun 16, 2025 β’ 278
Running Featured 1.42k FineWeb: decanting the web for the finest text data at scale π· 1.42k Explore and download the FineWeb webβscale text dataset