A 1B standard Transformer rivals Evo 2 40B on VEP Collection m5.1 release; scaling ladder; 5 training datasets; 5 validation probes; Mendelian/SGE evaluation only. Blog: https://openathena.ai/blog/marin-dna/ • 21 items • Updated Aug 6 • 5
Delphi Collection Marin's first open scaling suite. 88 base models, 3e18 → 1e23 FLOPs. https://openathena.ai/blog/delphi • 89 items • Updated May 19 • 12
Problems with Chinchilla Approach 2: Systematic Biases in IsoFLOP Parabola Fits Paper • 2603.22339 • Published Mar 21 • 5
Scaling Open Discrete Audio Foundation Models with Interleaved Semantic, Acoustic, and Text Tokens Paper • 2602.16687 • Published Feb 18 • 5
Olmo 3 Pre-training Collection All artifacts related to Olmo 3 pre-training • 10 items • Updated Dec 23, 2025 • 36
SynthesizeMe! Inducing Persona-Guided Prompts for Personalized Reward Models in LLMs Paper • 2506.05598 • Published Jun 5, 2025 • 8
Gemstone Models Collection Our 22 open source Gemstone models for scaling laws range from 50M to 2B parameters, spanning 11 widths from 256 to 3072 and 18 depths from 3 to 80. • 66 items • Updated Mar 2 • 11
AutoLibra: Agent Metric Induction from Open-Ended Feedback Paper • 2505.02820 • Published May 5, 2025 • 3
Nemotron-CC: Transforming Common Crawl into a Refined Long-Horizon Pretraining Dataset Paper • 2412.02595 • Published Dec 3, 2024 • 8
Mind the Gap! Static and Interactive Evaluations of Large Audio Models Paper • 2502.15919 • Published Feb 21, 2025 • 4
view article Article Optimizing Pretraining Data Mixes with LLM-Estimated Utility WillHeld • Jan 22, 2025 • 5
Tutor CoPilot: A Human-AI Approach for Scaling Real-Time Expertise Paper • 2410.03017 • Published Oct 3, 2024 • 29