AgentDish directory
verification
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#184
↓ -3
Copperhead
An open-source AI engineering platform that edits, documents, and verifies KiCad circuit boards from a prompt or brief, with gated stages for spec, architecture, parts, schematic, layout, outputs, firmware, and dev plan. |
Developer Tools / AI Development Tools | 88 | ↓ -3 | 2 days ago | Details |
|
#388
↓ -4
Collie
Local-first coding agent that runs on the user’s machine and can work across the browser, desktop, terminal, and editor surfaces. The repo says it verifies fixes by reproducing failures and re-running assertions, with install instructions for Windows, macOS, and Linux. |
Developer Tools / AI Coding Agents | 87 | ↓ -4 | 42 days ago | Details |
|
#570
↑ +2
verbatimeter
A Python package and CLI for checking how closely an AI-generated answer reuses its source text, with support for verbatim matching, subsequence matching, quotation verification, and pipeline gating. |
Developer Tools / Code Assistant | 86 | ↑ +2 | 62 days ago | Details |
|
#730
↓ -3
hwatu
A verification browser for AI coding agents built on WebKitGTK, aimed at fast visual checks, page assertions, and human handoff. The repository shows install steps, CLI usage, MCP support, and a concrete agent workflow. |
Developer Tools / AI Coding Tools | 85 | ↓ -3 | 47 days ago | Details |
|
#828
↓ -6
vise
A local CLI that records a codebase’s behavior and gates AI-driven refactors against a frozen baseline. It runs declared probes, compares outputs byte-for-byte, and returns a simple pass/fail signal for agent workflows. |
Developer Tools / AI DevOps / Code Quality | 84 | ↓ -6 | 7 days ago | Details |
|
#830
↓ -6
WorldCut
WorldCut is a TypeScript package that checks whether observations from independent systems satisfy declared version and time relationships before an autonomous agent makes a decision. |
Developer Tool / AI Agent Infrastructure | 84 | ↓ -6 | 8 days ago | Details |
|
#1155
↓ -3
evidence-verify
A Go/Python tool for verifying AI audit logs, checking chain integrity, tamper detection, and whether a log is independently anchored outside the producing system. |
Developer Tools / AI Compliance / Audit Logging | 83 | ↓ -3 | 33 days ago | Details |
|
#1266
↑ +194
Q2 2026 MCP Ecosystem Health
A research report on the current MCP ecosystem, with live crawl numbers, verification rates, category breakdowns, and examples of both strong and weak MCP-positive sites. |
Research / AI research | 83 | ↑ +194 | 129 days ago | Details |
|
#1337
↓ -2
Evidence-to-Skill
A GitHub repository that turns untrusted source material into compact AI skills using evidence gates, validation steps, and audit output. |
Developer Tools / AI Safety / Agent Skills | 82 | ↓ -2 | 40 days ago | Details |
|
#1353
↓ -2
Cinchor
Cinchor provides tamper-evident accountability records for AI agent actions, letting users verify a real decision against an on-chain record and test whether any field has been altered. |
Developer Tools / AI Observability | 82 | ↓ -2 | 50 days ago | Details |
|
#1382
↓ -2
Make No Mistakes
An open-source enforcement layer for AI coding agents that adds frozen specs, tamper-detected tests, independent verification, and hard-blocking gates so unverified work cannot pass. |
Developer Tools / AI Coding | 82 | ↓ -2 | 67 days ago | Details |
|
#1473
↑ +2
Fermat's Last Theorem in Lean 4
A Lean 4 repository containing a machine-checked proof of Fermat’s Last Theorem, with build verification details and browsable HTML documentation for the proof structure. |
Developer Tools / Formal Methods / Proof Verification | 81 | ↑ +2 | 6 days ago | Details |
|
#1504
↑ +2
AI Agent Safety & Alignment — Agent Bayes
An early-access AI research assistant that presents a cited mindmap of AI agent safety and alignment research, with expandable nodes, source-backed claims, and built-in verification checks. |
Developer Tools / Code Assistant | 81 | ↑ +2 | 72 days ago | Details |
|
#1629
↑ +2
Marker
Marker is an AI testing and verification platform for voice agents, with a focus on simulation, transcripts, judgments, and versioned evidence across the workflow. |
Developer Tools / AI Testing & Evaluation | 79 | ↑ +2 | 43 days ago | Details |
|
#2040
→ 0
Code as Agent Harness
A research page and preprint about using code as the runtime layer for agent systems, with a taxonomy of harness interfaces, harness mechanisms, and scaling patterns for multi-agent workflows. |
Developer Tools / Code Assistant | 71 | → 0 | 113 days ago | Details |