Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Spaces:
zaid646
/
llm-evaluator
like
0
Paused
App
Files
Files
Community
Fetching metadata from the HF Docker repository...
main
llm-evaluator
/
src
17.7 kB
Ctrl+K
Ctrl+K
1 contributor
History:
4 commits
zaid646
Add 2s delay between target and judge calls within each sample
edb4454
about 1 month ago
__init__.py
Safe
0 Bytes
Initial commit: LLM-as-a-Judge evaluator
about 1 month ago
config.py
Safe
493 Bytes
Fix target to call LLM directly (no sandbox dependency); reduce to 3 golden samples
about 1 month ago
dashboard.py
Safe
4.21 kB
Initial commit: LLM-as-a-Judge evaluator
about 1 month ago
ingest.py
Safe
794 Bytes
Initial commit: LLM-as-a-Judge evaluator
about 1 month ago
judger.py
Safe
3.19 kB
Add retry with exponential backoff for rate limits; add delay between samples
about 1 month ago
models.py
Safe
862 Bytes
Initial commit: LLM-as-a-Judge evaluator
about 1 month ago
pipeline.py
Safe
4.01 kB
Add 2s delay between target and judge calls within each sample
about 1 month ago
retry.py
Safe
916 Bytes
Add retry with exponential backoff for rate limits; add delay between samples
about 1 month ago
target.py
Safe
1.52 kB
Add retry with exponential backoff for rate limits; add delay between samples
about 1 month ago
telemetry.py
Safe
779 Bytes
Initial commit: LLM-as-a-Judge evaluator
about 1 month ago
websearch.py
Safe
906 Bytes
Initial commit: LLM-as-a-Judge evaluator
about 1 month ago