Commit History

Add submission checklist to README with links to environment, training scripts, runnable Colab, training logs, results, and story/writeup.
e29f215

ambuj.raj commited on

Report real rollout reward metrics.
601bcb2

ambuj.raj commited on

Add shared Colab link to README.
713640e

ambuj.raj commited on

Enhance README and demo app with improved model demo status reporting. Added a TL;DR section to README for quick understanding of SchemaQuake's functionality. Updated the run_gpu_model_demo function to return detailed model loading status and integrated a new function to assess booking status in the demo app.
7f31b27

ambuj.raj commited on

Harden Model Demo loading fallback.
6c04ff6

ambuj.raj commited on

Refactor model loading and action handling in SchemaQuake demo and training notebook. Updated imports to include new Action and Observation types. Enhanced logic for selecting model checkpoints and executing actions based on environment state. Improved JSON extraction method for better robustness. Adjusted logging to differentiate between proposed and executed actions.
abf6af8

ambuj.raj commited on

Enhance SchemaQuake documentation and training notebook. Updated Blog.md and README.md to clarify the RL environment aspect of SchemaQuake. Added supporting evidence for the training results in training-results.md. Improved the training notebook with clearer instructions and code for model loading and evaluation. Adjusted demo app to reflect changes in model loading and checkpoint handling.
81d5f5e

ambuj.raj commited on

Load trained model repo in demos.
6d8ac6f

ambuj.raj commited on

Fix Model Demo startup on Space.
2e96338

ambuj.raj commited on

Add GPU adapter demo paths.
eba709a

ambuj.raj commited on

Update SchemaQuake training notebook for clarity in documentation. Adjusted the description to enhance readability and ensure reproducibility in the training path using Hugging Face TRL.
0fc727d

ambuj.raj commited on

Update SchemaQuake training notebook and training output metrics.
8b11173

ambuj.raj commited on

Enhance evaluation metrics and documentation for imperfect heuristic agent.
0d3c5df

ambuj.raj commited on

Add inline numeric values to training results.
895f063

ambuj.raj commited on

Consolidate final submission documentation.
584a30a

ambuj.raj commited on

Rename mini-blog to Blog.
49944a8

ambuj.raj commited on

Use final result artifacts and enrich submission story.
d6a009c

ambuj.raj commited on

Use PyPI-available OpenEnv dependency for Space build.
7ad653e

ambuj.raj commited on

Add final submission checklist and training results docs.
0765ec0

ambuj.raj commited on

Prepare SchemaQuake final OpenEnv hackathon submission.
e8c01a7

ambuj.raj commited on

Upgrade SchemaQuake training quality pipeline for Qwen 3B.
576f8c1

ambuj.raj commited on

Add hard drift evaluation benchmark.
e8e3732

ambuj.raj commited on

Train with real SchemaQuake rollout rewards.
ae47a56

ambuj.raj commited on

Add live training logs and curve visualizations in Gradio.
174de56

ambuj.raj commited on

Enable true GRPO configuration on Space.
330b53d

ambuj.raj commited on

Fix GRPO training config and fallback behavior.
25288b3

ambuj.raj commited on

SchemaQuake HF deployment
3a95281

ambuj.raj commited on