Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Spaces:
ArunMendu
/
Med_mao
like
0
Running
App
Files
Files
Community
Fetching metadata from the HF Docker repository...
main
Med_mao
/
tests
/
eval
38.8 kB
Ctrl+K
Ctrl+K
1 contributor
History:
5 commits
Mendu9
feat: Tier 1-3 optimisations + 4-agent audit — streaming queue, BM25 JSON, IP rate limit, async council, JSON logging, eval feedback loop, PubMed fallback, multi-worker deploy
ee211b9
2 months ago
fixtures
fix: rebuild biomedical entity graph with scispaCy NER, clean bad nodes
2 months ago
golden
feat: eval harness, guardrails, graph explorer UI, multi-provider web search
3 months ago
__init__.py
Safe
0 Bytes
feat: add NLI entailment checker using cross-encoder/nli-deberta-v3-small
3 months ago
conftest.py
Safe
1.72 kB
fix: frontend import path, graph_explorer scale arg, routing test mocks
3 months ago
test_council_veto.py
Safe
1.46 kB
feat: eval harness, guardrails, graph explorer UI, multi-provider web search
3 months ago
test_golden_dataset.py
Safe
3.46 kB
feat: Tier 1-3 optimisations + 4-agent audit — streaming queue, BM25 JSON, IP rate limit, async council, JSON logging, eval feedback loop, PubMed fallback, multi-worker deploy
2 months ago
test_hallucination.py
Safe
969 Bytes
feat: eval harness, guardrails, graph explorer UI, multi-provider web search
3 months ago
test_nli_checker.py
Safe
1.6 kB
feat: add NLI entailment checker using cross-encoder/nli-deberta-v3-small
3 months ago
test_rag_recall.py
Safe
825 Bytes
feat: eval harness, guardrails, graph explorer UI, multi-provider web search
3 months ago
test_ragas_scores.py
Safe
1.09 kB
fix: rebuild biomedical entity graph with scispaCy NER, clean bad nodes
2 months ago
test_retrieval_metrics.py
Safe
9.5 kB
feat: Tier 1-3 optimisations + 4-agent audit — streaming queue, BM25 JSON, IP rate limit, async council, JSON logging, eval feedback loop, PubMed fallback, multi-worker deploy
2 months ago
test_routing.py
Safe
848 Bytes
feat: eval harness, guardrails, graph explorer UI, multi-provider web search
3 months ago