-
-
-
-
-
-
Inference Providers
Active filters: Reward
Text Classification
• 2B • Updated
• 3
• 2
mradermacher/SmolTulu-1.7b-RM-GGUF
2B • Updated
• 122
mradermacher/SmolTulu-1.7b-RM-i1-GGUF
2B • Updated
• 93
Teen-Different/squiral_maze
Reinforcement Learning
• Updated
Text Classification
• Updated
• 4
• 9
Text Classification
• Updated
• 7
• 1
Text Classification
• Updated
• 15
• 25
Text Classification
• Updated
• 13
• 5
wangclnlp/GRAM-RR-LLaMA-3.1-8B-RewardModel
Text Generation
• 8B • Updated
• 12
• 2
wangclnlp/GRAM-RR-LLaMA-3.2-3B-RewardModel
Text Generation
• 3B • Updated
• 4
mradermacher/GRAM-RR-LLaMA-3.2-3B-RewardModel-GGUF
3B • Updated
• 42
mradermacher/GRAM-RR-LLaMA-3.2-3B-RewardModel-i1-GGUF
3B • Updated
• 41
mradermacher/GRAM-RR-LLaMA-3.1-8B-RewardModel-GGUF
8B • Updated
• 29
• 1
mradermacher/GRAM-RR-LLaMA-3.1-8B-RewardModel-i1-GGUF
8B • Updated
• 176
• 1