Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Spaces:
below-threshold
/
ai-response-validator
like
0
Sleeping
App
Files
Files
Community
Fetching metadata from the HF Docker repository...
main
ai-response-validator
/
eval
70.1 kB
Ctrl+K
Ctrl+K
3 contributors
History:
8 commits
mbochniak01
Add parallel LangChain RAG engine; harden pipeline from architecture audit
4e2b267
16 days ago
bot-answers.json
Safe
13.3 kB
Faithfulness: mean sentence scoring, strip chunk title prefix, lower threshold to 0.35
3 months ago
calibrate.py
Safe
2.89 kB
Fix compat, bugs, and types; expand retail KB
3 months ago
compare_faithfulness.py
Safe
4.15 kB
Replace HHEM with sentence-level NLI, add claim decomposition and drift detection
3 months ago
compare_retrieval.py
Safe
3.38 kB
Add parallel LangChain RAG engine; harden pipeline from architecture audit
16 days ago
drift.py
Safe
6.34 kB
Replace HHEM with sentence-level NLI, add claim decomposition and drift detection
3 months ago
golden-dataset.yaml
Safe
17.9 kB
Address Gate 5 audit gaps
3 months ago
metrics.py
Safe
14.1 kB
Fix compat, bugs, and types; expand retail KB
3 months ago
simulate_traffic.py
Safe
8.01 kB
Replace HHEM with sentence-level NLI, add claim decomposition and drift detection
3 months ago