AgentDish directory

repository

Accepted listings with this tag.

Listing Category Score Trend Checked
#1699 ↑ +6
AGY Memory Engine

A lightweight SQLite FTS5 memory layer and MCP server for Google Antigravity agents, with CLI tools for storing, searching, and syncing facts.

Developer Tools / AI Memory / MCP 78 ↑ +6 21 days ago Details

A DeepSeek Harness plugin that uses an LLM to verify and rank candidate outputs with select, compare, track, and rollout tools.

Developer Tools / AI DevTools 78 ↑ +6 22 days ago Details
#1706 ↑ +6
aakit

Aakit is a developer tool for measuring and tracking assumptions made by coding agents, linking each assumption to supporting code and flagging which ones fail. The repo includes a clear explanation of the idea, prior art, and early experiment results.

Developer Tools / Code Assistant 78 ↑ +6 30 days ago Details

An open-source project for running a small LLM and limited agent workflows on ESP32-P4 hardware, with model details, architecture notes, examples, and setup instructions.

AI Developer Tool / Edge AI / Embedded LLM 78 ↑ +6 36 days ago Details
#1735 ↑ +6
LLM Tools

A Python tool manager for AI agents that uses an LLM-tools.txt registry to install, pin, describe, and execute tools with one predictable contract.

Developer Tools / AI Tooling 78 ↑ +6 73 days ago Details
#1737 ↑ +6
use-zerostack

A skill for coding agents that routes slash commands and trigger phrases to zerostack, a lightweight CLI coding agent for coding, planning, reviewing, and orchestration tasks.

Developer Tools / AI Agents 78 ↑ +6 74 days ago Details
#1748 ↑ +6
brain-map-skill

An agent skill and Python builder that turns Markdown note folders from Obsidian or gbrain into an interactive HTML knowledge map with a force graph, growth timeline, filters, and note detail panels. The repo includes a prebuilt demo for quick inspection and supports use in Claude Code, OpenAI Codex, Cursor, and simila

Developer Tools / AI Agents 78 ↑ +6 90 days ago Details
#1774 ↑ +6
MaragingLoop

An experimental autonomous bare-metal OS agent that uses an LLM, VirtualBox, and screenshot inspection to iterate on low-level code, boot a VM, and refine behavior from execution results.

Developer Tools / AI Agent Framework 78 ↑ +6 117 days ago Details

A metadata-first trust control plane for authorized security workflows, evidence retention, release trust, and business-flow proof. The repository includes role-based docs, quick-start commands, safety boundaries, and release-trust materials.

Security / Security Operations / Trust Infrastructure 78 ↓ -181 128 days ago Details

A GitHub example that audits LangChain’s RAG quickstart with retrieval-quality metrics, flags off-topic and out-of-distribution queries, and surfaces ranking and calibration issues with charts and results files.

Developer Tool / RAG Evaluation 78 ↑ +4 129 days ago Details
#1794 → 0
Claude Style Patch

A drop-in CLAUDE.md style section that changes how Claude writes prose. The repo includes a clear install path, the style rules themselves, and notes on when the patch works best or degrades.

Writing / Copywriting 77 → 0 3 days ago Details

Public evaluation code for an Agent Memory Leaderboard, including answer-generation and scoring contracts for comparing LLM agent memory systems.

Developer Tool / AI Evaluation / Benchmarking 77 → 0 40 days ago Details
#1802 → 0
agent-skills

A repository of Claude Code and Codex skills for evaluating and designing multi-agent systems, centered on an /agent-architect skill for auditing, reviewing, and designing agent setups.

Developer Tools / AI Development 77 → 0 41 days ago Details
#1819 → 0
Attestor

Attestor is a TypeScript control plane for high-risk AI-driven operations. It sits between an AI-generated request and the real system action, applying policy, approval, scope, freshness, replay, and evidence checks before returning admit, narrow, review, or block.

Developer Tools / AI Safety / Governance 77 → 0 83 days ago Details

A GitHub repo that teaches users how to write and debug evals through an agent-guided skill called learn-evals. It includes setup steps, exercises, examples, and a path to install the skill in other projects.

Developer Tool / AI Evaluation / Agent Skill 76 ↓ -1 just now Details

A GitHub repo for SHOVE, a turn-based CLI tactics puzzle created by Claude Fable 5 and iterated through six play sessions until it was judged fun. The page includes a brief, writeup, source code, logs, and the main game script.

Games / AI-generated game 76 ↓ -1 27 days ago Details
#1854 ↓ -1
CaLLMar

A prompt-based medieval fantasy text adventure for LLM chats, with character creation, combat, companions, inventory tracking, and a main quest to find four mysterious items.

Games / Interactive Fiction 76 ↓ -1 40 days ago Details
#1885 ↓ -1
ITB Engine

A research repository and local web app for testing quantum gravity theory-space exclusions against encoded consistency constraints. It includes a CLI, a localhost interface, test coverage, and an LLM-powered research agent for running searches and report generation.

Research / Scientific Computing 76 ↓ -1 123 days ago Details

A GitHub-hosted technical guide for running a multi-tier agent system on Termux, with SQLite used for messaging, scheduling, and audit logs, plus thermal-aware admission control for Android devices.

Developer Tools / AI Agent Infrastructure 75 → 0 19 days ago Details

A primer for non-coders to design and implement software through structured dialogue with an AI. It explains the CGP method, its goals, how it differs from orchestrator/worker patterns, and how to use the matching primer version.

AI/Prompting Method / AI-assisted software design 75 → 0 50 days ago Details
#1898 → 0
xtra

Python framework for conversational social engineering detection using a finite state machine. The README says it tracks turn-by-turn signals like flattery density, give/ask ratio, escalation velocity, and scope mismatch, and exposes a simple `Xtra().analyze(turns)` usage example.

Developer Tool / AI Safety / Security 75 → 0 77 days ago Details

A draft specification for an offline emergency dispatch protocol designed for on-device AI models on phones when no dispatcher is available.

AI safety / emergency response / On-device dispatch protocol 74 ↓ -1 22 days ago Details
#1927 ↓ -1
Jekyll-Hyde

A Hermes plugin that uses adversarial LLM clones to confront sandbagging and reward-hacking behavior during agent sessions.

Developer Tools / Code Assistant 74 ↓ -1 33 days ago Details
#1935 ↓ -1
Senbonzakura

Open-source tool for removing refusal behavior from open-weight transformer language models by identifying and orthogonalizing refusal directions in activation space.

Developer Tools / AI Model Editing / Alignment Research 74 ↓ -1 55 days ago Details