Inference Providers
Active filters: dapo
yannabadie/sage-topology-policy-v2
Reinforcement Learning
• Updated • 3
yannabadie/sage-topology-orchestrator
Text Generation
• Updated • 17
• 3
mradermacher/DAPO-Coding-Qwen2.5-1.5B-Instruct-GGUF
2B • Updated • 135
• 2
Text Generation
• 2B • Updated • 19
• 1
Text Generation
• 2B • Updated • 17
Text Generation
• 8B • Updated • 11
Text Generation
• 8B • Updated • 29
• Text Generation
• 8B • Updated • 25
Text Generation
• 8B • Updated • 32
• • 1
mradermacher/DAPO-No-DS-GGUF
2B • Updated • 108
mradermacher/MMR-DAPO-8B-GGUF
8B • Updated • 94
• 1
mradermacher/MMR-DAPO-7B-GGUF
8B • Updated • 24
mradermacher/DAPO-No-DS-8B-GGUF
8B • Updated • 115
mradermacher/DAPO-No-DS-7B-GGUF
8B • Updated • 185
Text Generation
• 2B • Updated • 19
Text Generation
• 8B • Updated • 12
• Text Generation
• 8B • Updated • 23
• 1
mradermacher/DAPO-7B-GGUF
8B • Updated • 306
• 1
mradermacher/DAPO-8B-GGUF
8B • Updated • 112
srallabandi0225/inframind-0.5b-grpo
Text Generation
• 0.5B • Updated • 15
• • 7
srallabandi0225/inframind-0.5b-dapo
Text Generation
• 0.5B • Updated • 7
AmirhoseinGH/Gnosis-Qwen3-1.7B-Hybrid
Text Classification
• 2B • Updated • 133
AmirhoseinGH/Gnosis-Qwen3-4B-Instruct-2507
Text Classification
• 4B • Updated • 24
AmirhoseinGH/Gnosis-Qwen3-4B-Thinking-2507
Text Classification
• 4B • Updated • 10
AmirhoseinGH/Gnosis-Qwen3-8B
Text Classification
• 8B • Updated • 24
mradermacher/inframind-0.5b-dapo-GGUF
Reinforcement Learning
• 0.5B • Updated • 115
mradermacher/inframind-0.5b-grpo-GGUF
Reinforcement Learning
• 0.5B • Updated • 71
kangdawei/MMR-Sigmoid-DAPO
Text Generation
• 2B • Updated • 13
kangdawei/MMR-Sigmoid-DAPO-7B
Text Generation
• 8B • Updated • 13