Edit Models filters
Model Tree
Apps
Inference Providers
One-click Deployment
Models
443
Active filters: multi-agent
36n9/Vehuiah-Draco-20260425_054044
36n9/Vehuiah-Draco-20260425_054127
36n9/Vehuiah-Draco-20260425_054202
36n9/Vehuiah-Draco-20260425_054238
36n9/Vehuiah-Draco-20260425_054312
36n9/Vehuiah-Draco-20260425_054347
36n9/Vehuiah-Draco-20260425_054423
36n9/Vehuiah-Draco-20260425_054459
pragunk/PropagationShield
Text Generation • 8B • Updated • 16
Bharath-1608/negotiation-agent-grpo
Updated
ujjwalpardeshi/chakravyuh-analyzer-lora-v2
Text Generation • Updated • 4
Steven668866/qwen3-8b-grpo-teaching-phase1
Updated • 6
Timusgeorge/SynthAudit-Qwen2.5-3B-GRPO
Text Generation • Updated • 3
Bharavi/rpoe-x-qwen-0.5b-grpo
Reinforcement Learning • 0.5B • Updated • 24
M134pra/neon-syndicate-qwen25-sft
Text Generation • 0.5B • Updated • 22
srikrish2004/sentinel-qwen3-4b-grpo
Text Generation • Updated • 1
IshikaMahadar/hiring-fleet-grpo-adapter
Text Generation • Updated
garvitsachdeva/spindleflow-rl
Reinforcement Learning • Updated • 1
Prathamesh0292/market-rl-stage1
Reinforcement Learning • Updated
helloAK96/chaosops-grpo-lora
Text Generation • Updated • 3
kartikraut09/ecocloud-grpo-qwen
Text Generation • 0.5B • Updated • 20
helloAK96/chaosops-grpo-lora-p2
Text Generation • Updated • 3
OnurDemircioglu/OmniGPT-355M-Instruct
0.4B • Updated • 3 • 1
132ragini/triage-wars-llm
Reinforcement Learning • Updated
helloAK96/chaosops-grpo-lora-p3a
Text Generation • Updated • 2
RavichandraNayakar/openenv-grpo-merged
Reinforcement Learning • 8B • Updated • 7
balarajr/triage-hospital-agent
Text Generation • 4B • Updated • 11
nothr/boardroom-grpo-lora-L2-best
Text Generation • Updated • 1
coliseum034/coliseum-defender-grpo-live
Reinforcement Learning • Updated • 2