Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
MoralGym
non-profit
Activity Feed
Follow
3
AI & ML interests
reinforcement-learning, fine-tuning, multi-agent, alignment, social-dilemmas, game-theory
Recent Activity
bickery
published
a model
2 days ago
moralgym/qwen3-8b-grpo-pd-deon-tft-200-v2-seed1-step200
bickery
published
a model
2 days ago
moralgym/qwen3-8b-sdpo-pdshch-deon-tft-200-seed1-step200
bickery
published
a model
2 days ago
moralgym/qwen3-8b-sdpo-pd-deon-tft-200-seed1-step200
View all activity
Team members
3
models
15
Sort: Recently updated
moralgym/qwen3-8b-grpo-pd-deon-tft-200-v2-seed1-step200
Text Generation
•
Updated
2 days ago
•
17
moralgym/qwen3-8b-sdpo-pd-deon-tft-200-seed1-step200
Text Generation
•
Updated
2 days ago
•
12
moralgym/qwen3-8b-sdpo-pdshch-deon-tft-200-seed1-step200
Text Generation
•
Updated
2 days ago
•
11
moralgym/qwen3-8b-pd-grpo-util-tft-step150
Text Generation
•
Updated
9 days ago
•
8
moralgym/gemma-3-12b-it-pd-grpo-deon-tft-step150
Text Generation
•
Updated
9 days ago
•
8
moralgym/qwen3-8b-pd-grpo-deon-tft-step180
Text Generation
•
Updated
9 days ago
•
9
moralgym/gemma-2-9b-it-pd-sdpo-deon-repair-gen-step120
Text Generation
•
Updated
9 days ago
•
4
moralgym/qwen3-8b-pd-sdpo-deon-repair-gen-step120
Text Generation
•
Updated
9 days ago
•
4
moralgym/qwen3-8b-pd-sdpo-deon-repair-gen-step110
Text Generation
•
Updated
9 days ago
•
9
moralgym/qwen3-8b-pd-sdpo-deon-repair-gen-step90
Text Generation
•
Updated
9 days ago
•
7
View 15 models
datasets
0
None public yet