NeoMME Collection Meet NeoMME: a family of 260M and 800M Multimodal-Native Multilingual Encoders • 12 items • Updated 4 days ago • 33
💧 LFM2.5 Collection Collection of post-trained and base LFM2.5 models. • 16 items • Updated Aug 4 • 229
view article Article Borealis — open data, code, weights recipe for training Audio LLM AlexWortega • May 25 • 17
Instella-MoE ✨ Collection Family of fully open 16B MoE LLM with 2.8B active params per token, trained on AMD Instinct™ MI300 & MI325 GPUs. https://arxiv.org/abs/2609.00791 • 6 items • Updated 14 days ago • 19
Weak-to-Strong Generalization via Direct On-Policy Distillation Paper • 2607.05394 • Published Jul 8 • 144
NLA Models Collection Natural Language Autoencoders across Llama, Gemma, and Qwen. Code: https://github.com/kitft/natural_language_autoencoders/ • 8 items • Updated May 6 • 18
NVFP4 Collection Dynamic Unsloth NVFP4 Quants. Faster and more accurate. • 11 items • Updated Aug 26 • 68
S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence Paper • 2606.20515 • Published Jun 18 • 42
PP-OCRv6 Collection From 1.5M to 34.5M Parameters, Surpassing Billion-Scale VLMs on OCR Tasks • 20 items • Updated Aug 14 • 114
SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer Paper • 2605.15178 • Published May 14 • 92
Mistral Small 4 Collection A state-of-the-art model, open-weight, with a granular Mixture-of-Experts architecture that fuses instruct, reasoning and agentic skills. • 3 items • Updated Mar 16 • 82