AgentDish directory
Top Picks
The listings that cleared review with the strongest evidence.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
A DeepSeek paper about DSpark, a full-stack codebase for training and evaluating speculative decoding algorithms to speed up LLM inference. |
Developer Tool / AI Infrastructure | 75 | → 0 | 75 days ago | Details |
|
#1898
→ 0
xtra
Python framework for conversational social engineering detection using a finite state machine. The README says it tracks turn-by-turn signals like flattery density, give/ask ratio, escalation velocity, and scope mismatch, and exposes a simple `Xtra().analyze(turns)` usage example. |
Developer Tool / AI Safety / Security | 75 | → 0 | 77 days ago | Details |
|
#1899
→ 0
Giskard
Giskard presents an AI red-teaming and continuous evaluation platform focused on finding hallucinations, security issues, and agent vulnerabilities. The page explains how it compares tools for agent-level testing, regression checks, and guardrail workflows. |
Security / AI Red Teaming | 75 | → 0 | 79 days ago | Details |
|
#1900
→ 0
Clayem
AI-assisted public adjuster service that analyzes property insurance policies, prepares demand packages, and connects policyholders with licensed adjusters to negotiate denied or underpaid claims. |
Legal / Insurance Claims | 75 | → 0 | 87 days ago | Details |
|
#1901
→ 0
Dead Internet Feed
A live feed of AI-generated posts and comment threads presented as a parody of online discourse. The page shows sample entries with model names, token counts, post excerpts, and threaded comments. |
AI-powered content feed / AI-generated social/news feed | 75 | → 0 | 91 days ago | Details |
|
#1902
→ 0
minecraft-builder-skill
A GitHub repository for a Minecraft skill that gives coding agents build-design abilities, JavaScript build recipe generation, vanilla .nbt export, and interactive previews before placing structures in-game. |
Developer Tools / AI Development Tools | 75 | → 0 | 95 days ago | Details |
|
#1903
→ 0
100cc
Open-source TypeScript project for building a minimal coding agent that can write and modify its own code, with setup instructions and example commands in the README. |
Developer Tool / AI Agent Framework | 75 | → 0 | 100 days ago | Details |
|
Docker blog post about a real AI coding agent failure and how Docker Sandboxes aim to contain destructive execution mistakes. |
Developer Tools / Code Assistant | 75 | → 0 | 102 days ago | Details |
|
#1905
→ 0
cartographer
A Claude Code skill for scoping fuzzy or domain-heavy software requests by building a short problem-theory before coding. |
Developer Tool / AI Coding Assistant Skill | 75 | → 0 | 104 days ago | Details |
|
A GitHub research project documenting a long-form, multi-model analysis of LLM behavior across Claude, Gemini, ChatGPT, and Grok. The repo includes an executive summary, screenplay, technical white paper, and archive of logs and chat records. |
AI Research / LLM Evaluation & Analysis | 75 | → 0 | 108 days ago | Details |
|
An article about using GitHub Agentic Workflows to scale documentation QA with a tiered AI setup, combining deterministic checks with LLM-assisted reviews, centralized workflows, and MCP-backed knowledge. |
Writing / Copywriting | 75 | → 0 | 123 days ago | Details |
|
#1908
↓ -22
doola MCP
doola MCP lets users form a U.S. LLC inside AI chat tools like Claude and Replit via MCP, with a conversational flow that covers account setup, package selection, payment, and post-payment formation steps. |
Developer Tools / Copywriting | 75 | ↓ -22 | 129 days ago | Details |
|
A Prompt One blog post about compiled workflow agents, design-time tool selection, and why runtime agents should not choose tools dynamically. |
Writing / Copywriting | 74 | ↓ -1 | 45 hours ago | Details |
|
A paper page about FrogNano, a 4B coding agent trained with RL on roughly 1,500 SWE environments using synthetic tasks and no distillation from a larger model. |
Agents / Coding Agent | 74 | ↓ -1 | 2 days ago | Details |
|
#1911
↓ -1
Project HydraFusion
GitHub’s research preview on multi-model orchestration for coding tasks. The page describes adaptive workflow selection and includes benchmark comparisons against the Opus 5 baseline with estimated cost reduction claims. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 6 days ago | Details |
|
#1912
↓ -1
AntiAgent
AntiAgent is a human-in-the-loop planning tool that uses AI to ask sharp questions, surface priorities, and turn messy thoughts into a step-by-step execution plan. |
AI Productivity / Decision Support | 74 | ↓ -1 | 7 days ago | Details |
|
#1913
↓ -1
Oconee Runtime
Oconee Runtime is an enterprise AI governance product focused on policy enforcement for browser AI and coding agents. The page explains how it evaluates proposed agent actions against policy and supports allow, warn, and block outcomes. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 8 days ago | Details |
|
#1914
↓ -1
Guess the AI Website Design Model
An interactive quiz from Raq.com that asks users to identify which AI model created each of ten default website designs, with links to the live pages after each reveal. |
Design / Creative Tools | 74 | ↓ -1 | 10 days ago | Details |
|
A field note about testing a coding agent fully offline on a laptop with Ollama and Qwen3.8 27B, including the setup, failure modes, and the workaround that made the project finish. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 10 days ago | Details |
|
#1916
↓ -1
Extrapolation Under an Exact Verifier
A research page describing a verified circle-packing result produced by a frozen language model with an external memory loop, with published trace, attempts archive, and repository code/data. |
AI Research / LLM Experiments | 74 | ↓ -1 | 11 days ago | Details |
|
#1917
↓ -1
THROTTLE
An interactive simulation about how inference providers ration access to stronger LLMs when demand spikes. Users can compare the default industry rule with their own policy and explore who gets the strong model. |
AI Product / LLM Routing / Inference Management | 74 | ↓ -1 | 14 days ago | Details |
|
#1918
↓ -1
LPU Lite
An educational project that explains and demonstrates a homemade language processing unit inspired by Groq’s LPU, with a focus on running a small Transformer model and visualizing the idea behind AI hardware. |
Developer Tools / AI Infrastructure | 74 | ↓ -1 | 17 days ago | Details |
|
A browser-playable rebuild of Future Crew’s 1993 Second Reality demo, presented as reconstructed from the original sources rather than via emulation or video. |
AI-powered product / Interactive demo / browser experience | 74 | ↓ -1 | 21 days ago | Details |
|
#1920
↓ -1
Offline Emergency Dispatch Protocol
A draft specification for an offline emergency dispatch protocol designed for on-device AI models on phones when no dispatcher is available. |
AI safety / emergency response / On-device dispatch protocol | 74 | ↓ -1 | 22 days ago | Details |