IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Prompt-seed101 Reinforcement Learning • 15B • Updated about 23 hours ago • 13
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Prompt-seed202 Reinforcement Learning • 15B • Updated about 23 hours ago • 15
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Ablation-Prompt-seed303 Reinforcement Learning • 15B • Updated about 23 hours ago • 13
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Minimalist-seed101 Reinforcement Learning • 15B • Updated about 23 hours ago • 13
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Minimalist-seed202 Reinforcement Learning • 15B • Updated about 23 hours ago • 14
IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Minimalist-seed303 Reinforcement Learning • 15B • Updated about 22 hours ago • 14