ARC 01-ai/Yi-Coder-9B-Chat Text Generation • 9B • Updated Sep 12, 2024 • 10.4k • 216 AI-MO/NuminaMath-7B-TIR Text Generation • 7B • Updated Aug 14, 2024 • 420 • 352
AIMO AI-MO/NuminaMath-7B-TIR Text Generation • 7B • Updated Aug 14, 2024 • 420 • 352 Running Agents 436 Reward Bench Leaderboard 📐 436 Explore and compare model scores on RewardBench benchmarks KTO: Model Alignment as Prospect Theoretic Optimization Paper • 2402.01306 • Published Feb 2, 2024 • 24
Running Agents 436 Reward Bench Leaderboard 📐 436 Explore and compare model scores on RewardBench benchmarks
KTO: Model Alignment as Prospect Theoretic Optimization Paper • 2402.01306 • Published Feb 2, 2024 • 24
Fater.ai mistralai/Mistral-7B-Instruct-v0.1 Text Generation • 7B • Updated Jul 24, 2025 • 173k • • 1.85k Quiet-STaR: Language Models Can Teach Themselves to Think Before Speaking Paper • 2403.09629 • Published Mar 14, 2024 • 81 Improve Mathematical Reasoning in Language Models by Automated Process Supervision Paper • 2406.06592 • Published Jun 5, 2024 • 27
Quiet-STaR: Language Models Can Teach Themselves to Think Before Speaking Paper • 2403.09629 • Published Mar 14, 2024 • 81
Improve Mathematical Reasoning in Language Models by Automated Process Supervision Paper • 2406.06592 • Published Jun 5, 2024 • 27
AIMO AI-MO/NuminaMath-7B-TIR Text Generation • 7B • Updated Aug 14, 2024 • 420 • 352 Running Agents 436 Reward Bench Leaderboard 📐 436 Explore and compare model scores on RewardBench benchmarks KTO: Model Alignment as Prospect Theoretic Optimization Paper • 2402.01306 • Published Feb 2, 2024 • 24
Running Agents 436 Reward Bench Leaderboard 📐 436 Explore and compare model scores on RewardBench benchmarks
KTO: Model Alignment as Prospect Theoretic Optimization Paper • 2402.01306 • Published Feb 2, 2024 • 24
ARC 01-ai/Yi-Coder-9B-Chat Text Generation • 9B • Updated Sep 12, 2024 • 10.4k • 216 AI-MO/NuminaMath-7B-TIR Text Generation • 7B • Updated Aug 14, 2024 • 420 • 352
Fater.ai mistralai/Mistral-7B-Instruct-v0.1 Text Generation • 7B • Updated Jul 24, 2025 • 173k • • 1.85k Quiet-STaR: Language Models Can Teach Themselves to Think Before Speaking Paper • 2403.09629 • Published Mar 14, 2024 • 81 Improve Mathematical Reasoning in Language Models by Automated Process Supervision Paper • 2406.06592 • Published Jun 5, 2024 • 27
Quiet-STaR: Language Models Can Teach Themselves to Think Before Speaking Paper • 2403.09629 • Published Mar 14, 2024 • 81
Improve Mathematical Reasoning in Language Models by Automated Process Supervision Paper • 2406.06592 • Published Jun 5, 2024 • 27