Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Spaces:
ddevMhrn
/
viveka-env
like
0
Sleeping
App
Files
Files
Community
Fetching metadata from the HF Docker repository...
main
viveka-env
/
eval
5.55 MB
Ctrl+K
Ctrl+K
2 contributors
History:
13 commits
ddevMhrn
Sync to local main (commit 1e292cf): post-hackathon blog updates
f39c634
verified
2 months ago
fixtures
feat(eval): add AQI evaluation scripts and reliability diagrams
3 months ago
plots
Polish Viveka blog with references and minor fixes
2 months ago
results
Sync to local main (commit 1e292cf): post-hackathon blog updates
2 months ago
_synthetic_reliability.png
154 kB
xet
feat(eval): add AQI evaluation scripts and reliability diagrams
3 months ago
aqi_delta.py
Safe
2.57 kB
feat(eval): add AQI evaluation scripts and reliability diagrams
3 months ago
aqi_probe.py
Safe
19.5 kB
feat(inference, train, eval): enhance user prompt structure with memory orchestration fields
3 months ago
aqi_scenario_probe.py
Safe
2.64 kB
feat(inference, train, eval): enhance user prompt structure with memory orchestration fields
3 months ago
baseline_claude_haiku.json
Safe
147 kB
Sync to local main (commit 1e292cf): post-hackathon blog updates
2 months ago
baseline_claude_haiku_OLD_sonnet_rerun.json
Safe
21.2 kB
Sync to local main (commit 1e292cf): post-hackathon blog updates
2 months ago
baseline_claude_sonnet.json
Safe
22.2 kB
feat(eval): add baseline evaluation files for Claude and GPT models
3 months ago
baseline_gpt5.2_3per_tier.json
Safe
51.9 kB
feat(eval): add baseline evaluation files for Claude and GPT models
3 months ago
baseline_gpt_4o_mini_per_tier1.json
Safe
42.8 kB
feat(eval): add baseline evaluation files for Claude and GPT models
3 months ago
baseline_gpt_4o_mini_per_tier21.json
Safe
30.3 kB
feat(eval): add baseline evaluation files for Claude and GPT models
3 months ago
baseline_gpt_5_2_per_tier1.json
Safe
67.1 kB
feat(eval): add baseline evaluation files for Claude and GPT models
3 months ago
capability_report.py
Safe
11.4 kB
Sync to local main (commit 1e292cf): post-hackathon blog updates
2 months ago
holdout_eval.py
Safe
13.8 kB
feat(docs): enhance README and PITCH with detailed Viveka overview
3 months ago
plot_combined_curves.py
Safe
7.65 kB
docs(readme): audit fixes + Llama 1B sealed eval + 3-architecture results
3 months ago
plot_leaderboard.py
Safe
6.32 kB
feat(submission): trained Llama-3B sealed eval + leaderboard PNG + Gradio fix
3 months ago
plot_loss_curves.py
Safe
5.52 kB
feat(submission): canonical OpenEnv layout + loss curves + Llama-3B baseline
3 months ago
probe_set.json
Safe
4.09 kB
feat(eval): add AQI evaluation scripts and reliability diagrams
3 months ago
reliability_diagram.py
Safe
14.6 kB
feat(eval): add AQI evaluation scripts and reliability diagrams
3 months ago
reward_curve.py
Safe
5.48 kB
feat(eval): add AQI evaluation scripts and reliability diagrams
3 months ago
test_aqi_synthetic.py
Safe
1.39 kB
feat(eval): add AQI evaluation scripts and reliability diagrams
3 months ago