AgentDish directory

Trending

Listings with strong scores and recent movement.

Listing Category Score Trend Checked

arXiv paper introducing ScientistOne, an autonomous research system built around a chain-of-evidence framework to keep claims traceable and audit research outputs for verifiability.

Research / AI Research Agents 74 ↓ -1 34 days ago Details
#1941 ↓ -1
Jekyll-Hyde

A Hermes plugin that uses adversarial LLM clones to confront sandbagging and reward-hacking behavior during agent sessions.

Developer Tools / Code Assistant 74 ↓ -1 34 days ago Details
#1942 ↓ -1
FlowChartCharter

Open-source Python project for execution-first multi-agent workflows built around YAML Charterfiles, runtime schema enforcement, and a boss-agent hierarchy. The README includes install commands, a short usage tour, and a small code example.

Developer Tools / AI Orchestration 74 ↓ -1 35 days ago Details

A playable text-adventure that teaches Claude Code through real prompts, skills, and agent workflows.

Developer Tools / Code Assistant 74 ↓ -1 38 days ago Details
#1944 ↓ -1
vite-bundle-antillm

A Vite plugin that obfuscates client-side code in a way meant to make LLMs refuse de-obfuscation attempts. The README shows the project’s purpose, a usage note, and an example of the transformed output.

Developer Tools / AI Safety / Prompt Protection 74 ↓ -1 38 days ago Details

A blog post about Idle Frontier, a clicker game about building a frontier language model in 1,000 days, and the author’s experience making it with Claude, Godot, and Aseprite.

AI-powered product / Game / Simulation 74 ↓ -1 40 days ago Details

A research article describing an empirical runtime-monitoring approach for multi-turn LLM agents, backed by 3,175 runs across four benchmarks. It also points to an open-source implementation, state-harness, with Rust/Python support, LangGraph and CrewAI adapters, CLI tooling, and OpenTelemetry export.

Developer Tools / Code Assistant 74 ↓ -1 40 days ago Details

JFrog Security Research investigates a batch of SQLite CVEs and argues many are AI-generated or unsupported by the source code and PoC testing. The post includes a comparison matrix, methodology, and detailed breakdown of several alleged vulnerabilities.

Research / Security Research 74 ↓ -1 40 days ago Details

A long-form essay about why AI coding software factories break down, with discussion of harness engineering, benchmark limits, and safer ways to use coding agents.

Writing / Copywriting 74 ↓ -1 51 days ago Details
#1949 ↓ -1
Senbonzakura

Open-source tool for removing refusal behavior from open-weight transformer language models by identifying and orthogonalizing refusal directions in activation space.

Developer Tools / AI Model Editing / Alignment Research 74 ↓ -1 56 days ago Details

A RequestRocket article mapping the AI agent access-control stack across authentication, authorization, governance, identity, gateways, and runtime egress control. It explains how the categories differ, when each fits, and where RequestRocket sits in the landscape.

Developer Tools / Code Assistant 74 ↓ -1 61 days ago Details
#1951 ↓ -1
Monogram

Monogram is an AI app that turns prompts into an interactive visual interface instead of plain text. The page shows a live product pitch, demo entry points, and example use cases like recipes, movies, birthday planning, restaurant search, and EV comparisons.

AI Product / Visual Interface 74 ↓ -1 65 days ago Details

An in-depth Medium post from IronBee about how a verification loop affects AI coding agents, using Web-Bench and comparing DeepSeek with Claude Opus on a real web app task.

Developer Tools / Code Assistant 74 ↓ -1 67 days ago Details

arXiv paper on a cache-merging method for multi-agent latent reasoning, framing KV-cache composition as a convergent replicated state with deterministic merging.

Research / AI Research Paper 74 ↓ -1 71 days ago Details

A live, data-driven dashboard that tracks whether AI expansion is still intensifying or starting to cool, using infrastructure spend, data-center power, sovereign commitments, and supply-chain bottlenecks.

AI Data & Analytics / Market Intelligence 74 ↓ -1 71 days ago Details
#1955 ↓ -1
GENIUS AI Detector

A free browser-based tool that claims to detect whether text is AI-generated, with example inputs, inline explanations, and FAQ notes about privacy and local processing.

AI Tools / AI Detection 74 ↓ -1 76 days ago Details
#1956 ↓ -1
local-char

Open-source local AI character platform that runs with llama.cpp and offers CLI, TUI, and a simple web server for character roleplay without cloud accounts or subscriptions.

Developer Tools / AI Agents 74 ↓ -1 78 days ago Details
#1957 ↓ -1
Agent Departures

A web app that generates product or project ideas by pairing familiar tools with an added agent layer. The page frames it as an idea generator for “copy idea” style prompts, with a simple call to action to pull the lever for a new departure.

Writing / Copywriting 74 ↓ -1 83 days ago Details

A macOS terminal workflow that uses local speech-to-text plus the Pi coding agent to turn spoken requests into shell commands or terminal answers.

Developer Tools / CLI 74 ↓ -1 86 days ago Details

Anthropic research report on how people use Claude Code in practice, based on a large privacy-preserving analysis of session data. It covers task types, division of labor between user and model, and how domain expertise affects outcomes.

Developer Tools / Code Assistant 74 ↓ -1 86 days ago Details

A blog post describing an evaluation harness for comparing agentic coding tools and prompts across realistic SWE tasks, with token-cost results and model-specific behavior notes.

Developer Tools / Code Assistant 74 ↓ -1 90 days ago Details
#1961 ↓ -1
Meadow Mind

An open-source Python project that uses a local 7B diffusion language model to make real-time Gymnasium decisions with zero training. The README shows install steps, a code example, and benchmark-style results across several environments.

AI / Reinforcement Learning 74 ↓ -1 93 days ago Details

A shareable report for AI coding usage that summarizes how you work with tools like Claude Code, Codex, and Cursor, with sign-in via GitHub and an option to get printed copies.

Developer Tools / Code Assistant 74 ↓ -1 93 days ago Details
#1963 ↓ -1
LEVI

LEVI is a harness-first evolutionary framework for code and prompt optimization. It focuses on reducing LLM cost with diversity-preserving search, role-aware model routing, and a proxy benchmark, and presents comparative results against several existing systems.

Developer Tools / Code Assistant 74 ↓ -1 96 days ago Details