Reward Modeling Datasets
updated
Viewer
• Updated
• 37.1k • 4.55k
• 247
Viewer
• Updated
• 169k • 17.5k
• 1.67k
Viewer
• Updated
• 386k • 2.52k
• 323
PKU-Alignment/PKU-SafeRLHF
Viewer
• Updated
• 164k • 9.93k
• 178
openai/webgpt_comparisons
Viewer
• Updated
• 19.6k • 546
• 240
openai/summarize_from_feedback
Viewer
• Updated
• 194k • 5.67k
• 217
HuggingFaceH4/ultrafeedback_binarized
Viewer
• Updated
• 187k • 5.56k
• 323
Viewer
• Updated
• 183k • 1.13k
• 295
HuggingFaceH4/stack-exchange-preferences
Viewer
• Updated
• 10.8M • 4.6k
• 133
HuggingFaceH4/hhh_alignment
Viewer
• Updated
• 221 • 164
• 22
Birchlabs/openai-prm800k-stepwise-critic
Viewer
• Updated
• 1.09M • 98
• 45
prometheus-eval/Feedback-Collection
Viewer
• Updated
• 100k • 410
• 118
argilla/OpenHermesPreferences
Viewer
• Updated
• 989k • 567
• 211
Viewer
• Updated
• 8.11k • 6.07k
• 105
Viewer
• Updated
• 21.4k • 16.3k
• 439
Magpie-Align/Magpie-Pro-DPO-200K
Viewer
• Updated
• 207k • 8
• 7
argilla/magpie-ultra-v0.1
Viewer
• Updated
• 50k • 328
• 221