AgentDish directory
hallucination detection
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
A research repo for reducing hallucinations in LLM-generated code using semantic triangulation, with setup, benchmarking, experimentation, and reproducibility instructions. |
Developer Tool / Code Quality | 86 | ↑ +2 | 23 days ago | Details |
|
#504
↑ +2
verbatimeter
A Python package and CLI for checking how closely an AI-generated answer reuses its source text, with support for verbatim matching, subsequence matching, quotation verification, and pipeline gating. |
Developer Tools / Code Assistant | 86 | ↑ +2 | 52 days ago | Details |
|
#731
↓ -6
hedgemony
Hedgemony is a Python tool that checks AI-written code for fabricated packages, invented APIs, impossible calls, and contradictions between code and stated examples, using interpreter and registry checks rather than another model. |
Developer Tools / Code Analysis | 84 | ↓ -6 | 27 hours ago | Details |
|
#1596
↑ +6
Verified RAG: every sentence checked
A blog post about verifiable RAG that benchmarks open-source NLI verifiers against Claude on RAGTruth and describes a Python library for sentence-level citation and claim verification. |
AI / RAG / Verification & Hallucination Detection | 78 | ↑ +6 | 91 days ago | Details |
|
#1729
→ 0
Giskard
Giskard presents an AI red-teaming and continuous evaluation platform focused on finding hallucinations, security issues, and agent vulnerabilities. The page explains how it compares tools for agent-level testing, regression checks, and guardrail workflows. |
Security / AI Red Teaming | 75 | → 0 | 69 days ago | Details |