Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

Spaces:
eliezeravihail
/
torchdocs-agent
Running

App Files Files Community
Fetching metadata from the HF Docker repository...
torchdocs-agent / tests /eval
11.9 kB
Ctrl+K
Ctrl+K
  • 4 contributors
History: 10 commits
eliezer avihail
Answer-quality eval (LLM-judge) + hy3-led OpenRouter default (#81)
0942768 unverified 23 days ago
  • __init__.py
    0 Bytes
    Scaffold repo structure per M0 (pyproject, ruff, pytest, pre-commit) 29 days ago
  • test_checks.py
    2.7 kB
    Codebase quality pass: kill silent-truncation hacks, fix latent bugs, cover them (#46) 25 days ago
  • test_run_agentic.py
    1.5 kB
    Add the agentic benchmark: coverage via citations, agentic vs single-shot 23 days ago
  • test_run_judge.py
    3.85 kB
    Answer-quality eval (LLM-judge) + hy3-led OpenRouter default (#81) 23 days ago
  • test_run_retrieval.py
    2.92 kB
    Add the v1 eval dataset: 100 valid + 100 invalid + 20 agentic 23 days ago
  • test_run_v0.py
    932 Bytes
    Fix LLM transport drift, wire live static checks, clean config (#39) 25 days ago