Commit History
fix: include context_type and policy_mode in episode history entries 402a8c7
feat: add grader classes for all 3 DRL tasks 40f1fb9
feat: implement state interface, standardize schema IDs, and add task discovery endpoint a447d83
Gamucopia-Creatives commited on
fix: set default value for ResetRequest in reset_env endpoint to prevent validation errors 46d41dd
Gamucopia-Creatives commited on
feat: implement reinforced human memory for inference and enhance UI with audit capabilities and task labeling 85218c9
Gamucopia-Creatives commited on
refactor: normalize reward range to [0.01, 0.99] and standardize episode scoring key to score across environment and inference logic da09194
Gamucopia-Creatives commited on
feat: enhance safety moderation with keyword-based fallback, robust LLM response parsing, and unified UI insight formatting. 8ddb260
Gamucopia-Creatives commited on
refactor: update submission validator with strict log format checks, HF_TOKEN safety verification, and add FastAPI server for interactive task testing 1191a4e
Gamucopia-Creatives commited on