AgentDish directory
hallucination detection
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
A research repo for reducing hallucinations in LLM-generated code using semantic triangulation, with setup, benchmarking, experimentation, and reproducibility instructions. |
Developer Tool / Code Quality | 86 | ↑ +2 | 33 days ago | Details |
|
#570
↑ +2
verbatimeter
A Python package and CLI for checking how closely an AI-generated answer reuses its source text, with support for verbatim matching, subsequence matching, quotation verification, and pipeline gating. |
Developer Tools / Code Assistant | 86 | ↑ +2 | 62 days ago | Details |
|
#835
↓ -6
hedgemony
Hedgemony is a Python tool that checks AI-written code for fabricated packages, invented APIs, impossible calls, and contradictions between code and stated examples, using interpreter and registry checks rather than another model. |
Developer Tools / Code Analysis | 84 | ↓ -6 | 11 days ago | Details |
|
#1268
↓ -2
Spanda
Open-source LLM uncertainty and hallucination detection tool that estimates epistemic risk in microseconds using exact-match normalized entropy, with Rust gateway and Python package support. |
Developer Tools / Code Assistant | 82 | ↓ -2 | just now | Details |
|
#1756
↑ +6
Verified RAG: every sentence checked
A blog post about verifiable RAG that benchmarks open-source NLI verifiers against Claude on RAGTruth and describes a Python library for sentence-level citation and claim verification. |
AI / RAG / Verification & Hallucination Detection | 78 | ↑ +6 | 101 days ago | Details |
|
#1899
→ 0
Giskard
Giskard presents an AI red-teaming and continuous evaluation platform focused on finding hallucinations, security issues, and agent vulnerabilities. The page explains how it compares tools for agent-level testing, regression checks, and guardrail workflows. |
Security / AI Red Teaming | 75 | → 0 | 79 days ago | Details |