Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Spaces:
DeepikaChintamreddy
/
config-debug-env
like
0
Sleeping
App
Files
Files
Community
Fetching metadata from the HF Docker repository...
main
config-debug-env
644 kB
Ctrl+K
Ctrl+K
3 contributors
History:
50 commits
Deepikachintamreddy
fix: eliminate all 0.0 reward values across all files
bc98e17
4 months ago
server
fix: eliminate all 0.0 reward values across all files
4 months ago
.dockerignore
56 Bytes
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
.gitattributes
0 Bytes
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
.gitignore
19 Bytes
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
Dockerfile
276 Bytes
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
README.md
8.64 kB
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
__init__.py
187 Bytes
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
client.py
1.03 kB
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
inference.py
7.4 kB
fix: correct step format, per-task output, tuple grader returns
4 months ago
openenv.yaml
1.59 kB
fix: class-based graders with grade() method per Meta instructions
4 months ago
pyproject.toml
446 Bytes
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
requirements.txt
72 Bytes
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
test_env.py
540 Bytes
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
test_grader_audit.py
3.14 kB
CRITICAL FIX: Graders now return FLOAT ONLY (not tuples) for validator contract compliance - extracted float from raw grader tuples
5 months ago
test_k8s_multistep.py
1.77 kB
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
test_nginx_multistep.py
1.53 kB
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
test_task_graders.py
2.95 kB
Fix: Enforce strict open interval (0.01, 0.99) for all grader rewards - validator compliance
5 months ago
uv.lock
538 kB
Phase 2: Benchmark Quality Upgrades - Progressive graders, realistic multi-bug tasks, ground_truth exposure
5 months ago
verify_graders.py
1.61 kB
fix: resolve merge conflict in grader_api.py
4 months ago