AgentDish directory
Trending
Listings with strong scores and recent movement.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
arXiv paper introducing ScientistOne, an autonomous research system built around a chain-of-evidence framework to keep claims traceable and audit research outputs for verifiability. |
Research / AI Research Agents | 74 | ↓ -1 | 34 days ago | Details |
|
#1941
↓ -1
Jekyll-Hyde
A Hermes plugin that uses adversarial LLM clones to confront sandbagging and reward-hacking behavior during agent sessions. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 34 days ago | Details |
|
#1942
↓ -1
FlowChartCharter
Open-source Python project for execution-first multi-agent workflows built around YAML Charterfiles, runtime schema enforcement, and a boss-agent hierarchy. The README includes install commands, a short usage tour, and a small code example. |
Developer Tools / AI Orchestration | 74 | ↓ -1 | 35 days ago | Details |
|
#1943
↓ -1
MERIDIAN — A Claude Copilot Adventure
A playable text-adventure that teaches Claude Code through real prompts, skills, and agent workflows. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 38 days ago | Details |
|
#1944
↓ -1
vite-bundle-antillm
A Vite plugin that obfuscates client-side code in a way meant to make LLMs refuse de-obfuscation attempts. The README shows the project’s purpose, a usage note, and an example of the transformed output. |
Developer Tools / AI Safety / Prompt Protection | 74 | ↓ -1 | 38 days ago | Details |
|
A blog post about Idle Frontier, a clicker game about building a frontier language model in 1,000 days, and the author’s experience making it with Claude, Godot, and Aseprite. |
AI-powered product / Game / Simulation | 74 | ↓ -1 | 40 days ago | Details |
|
A research article describing an empirical runtime-monitoring approach for multi-turn LLM agents, backed by 3,175 runs across four benchmarks. It also points to an open-source implementation, state-harness, with Rust/Python support, LangGraph and CrewAI adapters, CLI tooling, and OpenTelemetry export. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 40 days ago | Details |
|
#1947
↓ -1
SQLite Critical CVEs or LLM Slop?
JFrog Security Research investigates a batch of SQLite CVEs and argues many are AI-generated or unsupported by the source code and PoC testing. The post includes a comparison matrix, methodology, and detailed breakdown of several alleged vulnerabilities. |
Research / Security Research | 74 | ↓ -1 | 40 days ago | Details |
|
#1948
↓ -1
Why Software Factories Fail
A long-form essay about why AI coding software factories break down, with discussion of harness engineering, benchmark limits, and safer ways to use coding agents. |
Writing / Copywriting | 74 | ↓ -1 | 51 days ago | Details |
|
#1949
↓ -1
Senbonzakura
Open-source tool for removing refusal behavior from open-weight transformer language models by identifying and orthogonalizing refusal directions in activation space. |
Developer Tools / AI Model Editing / Alignment Research | 74 | ↓ -1 | 56 days ago | Details |
|
A RequestRocket article mapping the AI agent access-control stack across authentication, authorization, governance, identity, gateways, and runtime egress control. It explains how the categories differ, when each fits, and where RequestRocket sits in the landscape. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 61 days ago | Details |
|
#1951
↓ -1
Monogram
Monogram is an AI app that turns prompts into an interactive visual interface instead of plain text. The page shows a live product pitch, demo entry points, and example use cases like recipes, movies, birthday planning, restaurant search, and EV comparisons. |
AI Product / Visual Interface | 74 | ↓ -1 | 65 days ago | Details |
|
An in-depth Medium post from IronBee about how a verification loop affects AI coding agents, using Web-Bench and comparing DeepSeek with Claude Opus on a real web app task. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 67 days ago | Details |
|
arXiv paper on a cache-merging method for multi-agent latent reasoning, framing KV-cache composition as a convergent replicated state with deterministic merging. |
Research / AI Research Paper | 74 | ↓ -1 | 71 days ago | Details |
|
#1954
↓ -1
Is the AI hype cooling down?
A live, data-driven dashboard that tracks whether AI expansion is still intensifying or starting to cool, using infrastructure spend, data-center power, sovereign commitments, and supply-chain bottlenecks. |
AI Data & Analytics / Market Intelligence | 74 | ↓ -1 | 71 days ago | Details |
|
#1955
↓ -1
GENIUS AI Detector
A free browser-based tool that claims to detect whether text is AI-generated, with example inputs, inline explanations, and FAQ notes about privacy and local processing. |
AI Tools / AI Detection | 74 | ↓ -1 | 76 days ago | Details |
|
#1956
↓ -1
local-char
Open-source local AI character platform that runs with llama.cpp and offers CLI, TUI, and a simple web server for character roleplay without cloud accounts or subscriptions. |
Developer Tools / AI Agents | 74 | ↓ -1 | 78 days ago | Details |
|
#1957
↓ -1
Agent Departures
A web app that generates product or project ideas by pairing familiar tools with an added agent layer. The page frames it as an idea generator for “copy idea” style prompts, with a simple call to action to pull the lever for a new departure. |
Writing / Copywriting | 74 | ↓ -1 | 83 days ago | Details |
|
A macOS terminal workflow that uses local speech-to-text plus the Pi coding agent to turn spoken requests into shell commands or terminal answers. |
Developer Tools / CLI | 74 | ↓ -1 | 86 days ago | Details |
|
Anthropic research report on how people use Claude Code in practice, based on a large privacy-preserving analysis of session data. It covers task types, division of labor between user and model, and how domain expertise affects outcomes. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 86 days ago | Details |
|
#1960
↓ -1
Evaluate Your Agentic Tooling
A blog post describing an evaluation harness for comparing agentic coding tools and prompts across realistic SWE tasks, with token-cost results and model-specific behavior notes. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 90 days ago | Details |
|
#1961
↓ -1
Meadow Mind
An open-source Python project that uses a local 7B diffusion language model to make real-time Gymnasium decisions with zero training. The README shows install steps, a code example, and benchmark-style results across several environments. |
AI / Reinforcement Learning | 74 | ↓ -1 | 93 days ago | Details |
|
#1962
↓ -1
Entelligence AI Wrapped 2026
A shareable report for AI coding usage that summarizes how you work with tools like Claude Code, Codex, and Cursor, with sign-in via GitHub and an option to get printed copies. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 93 days ago | Details |
|
#1963
↓ -1
LEVI
LEVI is a harness-first evolutionary framework for code and prompt optimization. It focuses on reducing LLM cost with diversity-preserving search, role-aware model routing, and a proxy benchmark, and presents comparative results against several existing systems. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 96 days ago | Details |