Inference Providers
Active filters: RLHF
NousResearch/Hermes-2-Pro-Llama-3-8B
Text Generation
• 8B • Updated • 221k
• • 446
OpenAssistant/reward-model-deberta-v3-large-v2
Text Classification
• Updated • 28.9k
• • 245
NousResearch/Hermes-2-Pro-Mistral-7B
Text Generation
• 7B • Updated • 3.39k
• 500
mlx-community/Hermes-2-Pro-Mistral-7B-4bit
1B • Updated • 93
• 4
mlx-community/Hermes-2-Pro-Mistral-7B-8bit
2B • Updated • 197
• 8
ruslanmv/Medical-Llama3-8B
Text Generation
• 8B • Updated • 1.17k
• • 106
mlx-community/Hermes-2-Pro-Mistral-7B-3bit
0.9B • Updated • 36
• 1
OpenAssistant/reward-model-deberta-v3-base
Text Classification
• Updated • 1.63k
• • 13
OpenAssistant/reward-model-electra-large-discriminator
Text Classification
• Updated • 440
• 5
OpenAssistant/reward-model-deberta-v3-large
Text Classification
• Updated • 967
• 26
Text Ranking
• 0.4B • Updated • 13
• 3
nicholasKluge/RewardModelPT
Text Classification
• 0.1B • Updated • 22
nicholasKluge/RewardModel
Text Classification
• 0.1B • Updated • 320
• 1
fb700/chatglm-fitness-RLHF
Updated • 268
fb700/Bofan-chatglm-Best-lora
Updated • 9
• 11
kubernetes-bad/Ligma-L2-13b
Updated • 9
• 3
Text Generation
• Updated • 308
• 205
berkeley-nest/Starling-LM-7B-alpha
Text Generation
• 7B • Updated • 1.87k
• 559
berkeley-nest/Starling-RM-7B-alpha
Updated • 222
• 104
LoneStriker/Starling-LM-7B-alpha-3.0bpw-h6-exl2
Text Generation
• Updated • 2
LoneStriker/Starling-LM-7B-alpha-4.0bpw-h6-exl2
Text Generation
• Updated • 5
• 1
LoneStriker/Starling-LM-7B-alpha-5.0bpw-h6-exl2
Text Generation
• Updated • 5
• 2
LoneStriker/Starling-LM-7B-alpha-6.0bpw-h6-exl2
Text Generation
• Updated • 4
• 1
LoneStriker/Starling-LM-7B-alpha-8.0bpw-h8-exl2
Text Generation
• Updated • 5
• 2
TheBloke/Starling-LM-7B-alpha-GGUF
7B • Updated • 1.76k
• 94
TheBloke/Starling-LM-7B-alpha-AWQ
Text Generation
• 7B • Updated • 16
• 9
second-state/Starling-LM-7B-alpha-GGUF
Text Generation
• 7B • Updated • 107
• 3
TheBloke/Starling-LM-7B-alpha-GPTQ
Text Generation
• 7B • Updated • 14
• 10
bartowski/Starling-LM-7B-alpha-old-exl2
Text Generation
• Updated tastypear/chatglm-fitness-RLHF-GGML