view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 29 days ago • 143
view article Article BenchMIRT: What are LLM benchmarks actually measuring? allenai • 30 days ago • 27
🌊 MoM 1.0 Collection long context models for MoM multilingual embeddings • 5 items • Updated 18 days ago • 4
Nemotron Supervised Fine-Tuning Collection SFT datasets covering math, code, chat, safety, agentic, VLM, multilingual, and specialized domains. • 44 items • Updated Aug 11 • 22
Mage Collection A family of lightweight multimodal models, including understanding and generation. • 8 items • Updated Jul 26 • 30
view article Article Ulysses Sequence Parallelism: Training with Million-Token Contexts kashif, stas • Mar 9 • 33
Luciole LLM Collection Open Source LLM in French, English, German, Spanish, Italian, Portuguese, Dutch and Arabic • 20 items • Updated 3 days ago • 14
Latxa Instruct Collection Instructing Large Language Models for Low-Resource Languages: A Systematic Study for Basque • 17 items • Updated Jun 25 • 2
view article Article Introducing BERTopic Integration with the Hugging Face Hub MaartenGr, davanstrien • May 31, 2023 • 12
NVIDIA Nemotron v3 Collection Open, Production-ready Enterprise Models • 33 items • Updated Aug 14 • 378