Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Spaces:
evaleval
/
general-eval-card
like
5
Running
App
Files
Files
Community
8
Fetching metadata from the HF Docker repository...
main
general-eval-card
/
lib
476 kB
Ctrl+K
Ctrl+K
8 contributors
History:
124 commits
j-chim
Unit-aware score display: units outrank the magnitude heuristic
0ab5b2a
2 days ago
backend-artifacts.ts
Safe
19.6 kB
Resolve score scales from producer-stamped registry bounds
3 days ago
benchmark-metadata-utils.ts
Safe
1.47 kB
Fix RewardBench2 key normalization for matrix leaderboard routing
4 months ago
benchmark-metadata.ts
Safe
1.65 kB
Differentiate audience modes and tighten eval navigation
4 months ago
benchmark-schema.ts
Safe
14.5 kB
Overlaps tab v2: single-source rows, per-source detail, slice-title contract
2 months ago
benchmark-tags.ts
Safe
11.1 kB
Update to uniform US spelling
about 1 month ago
cached-json-response.ts
Safe
3.11 kB
Cache + gzip the /models and /evals index payloads; remove temp diag route
3 months ago
clean-hierarchy.ts
Safe
71 kB
Fix validated evaluator displays
2 months ago
dashboard-data-client.ts
Safe
4.35 kB
Add merged-benchmark accessor + API routing (spec F1)
3 days ago
data-backend.ts
Safe
4.61 kB
Add merged-benchmark accessor + API routing (spec F1)
3 days ago
distribution-series.ts
Safe
5.68 kB
Fix embeds to filter by canonical benchmark only (no splits in embed until we decide how to display splits) + suppress popup on embed routes
2 months ago
duckdb.ts
Safe
4.21 kB
Add merged-benchmark accessor + API routing (spec F1)
3 days ago
eval-processing.ts
Safe
15.1 kB
Merged-page review fixes: exclude flagged rows, wire comparability, thread card
3 days ago
evaluator-logo.ts
Safe
2.18 kB
Fix model comparison + add evaluator logo
2 months ago
evaluators.ts
Safe
10.4 kB
Fix model comparison + add evaluator logo
2 months ago
glossary.ts
Safe
5.19 kB
Update wording
2 months ago
hf-data.ts
Safe
20.7 kB
Overlaps tab v2: single-source rows, per-source detail, slice-title contract
2 months ago
hierarchy-lookup.ts
Safe
5.59 kB
WIP: v2 cleanup checkpoint before merging origin/main
3 months ago
known-developers.ts
Safe
2.15 kB
Update about (#8)
2 months ago
known-issues.ts
Safe
1.67 kB
Separate policy and researcher views
4 months ago
merged-adapter.ts
Safe
7.87 kB
Close re-review residuals: slice-grain card fallback, shown-pool model count
3 days ago
metric-labels.ts
Safe
1.11 kB
Embeds: histogram route, leaderboard slices + sort, brand mark; cross-source row dedup
3 months ago
model-url-redirects-build.ts
Safe
5 kB
Update about (#8)
2 months ago
model-url-redirects.ts
Safe
105 kB
Regenerate model URL redirects for the 2026-08-13 snapshot
6 days ago
na-utils.ts
Safe
1.73 kB
WIP: v2 cleanup checkpoint before merging origin/main
3 months ago
overlaps.ts
Safe
22.2 kB
Unit-aware score display: units outrank the magnitude heuristic
2 days ago
param-range.ts
Safe
3.59 kB
Tighten eval cards UI and clean up stale local data
4 months ago
policy-summaries.ts
Safe
19.4 kB
Update wording
2 months ago
safe-storage.ts
Safe
1.32 kB
Tolerate missing localStorage in sandboxed embed iframes
about 2 months ago
score-scale.ts
Safe
11.2 kB
Unit-aware score display: units outrank the magnitude heuristic
2 days ago
sidecars.ts
Safe
15 kB
Fix model comparison + add evaluator logo
2 months ago
survey-content.ts
Safe
7.05 kB
Add survey submission and update survey text for public use
4 months ago
tutorials.ts
Safe
6.48 kB
Help docs: verified-checkmark gif + refreshed cross-post tutorial
2 months ago
utils.ts
Safe
6.93 kB
Add merged-benchmark accessor + API routing (spec F1)
3 days ago
view-data.ts
Safe
57.5 kB
Close re-review residuals: slice-grain card fallback, shown-pool model count
3 days ago