AgentDish directory

Top Picks

The listings that cleared review with the strongest evidence.

Listing Category Score Trend Checked

Anthropic research report on how people use Claude Code in practice, based on a large privacy-preserving analysis of session data. It covers task types, division of labor between user and model, and how domain expertise affects outcomes.

Developer Tools / Code Assistant 74 ↓ -1 85 days ago Details

A blog post describing an evaluation harness for comparing agentic coding tools and prompts across realistic SWE tasks, with token-cost results and model-specific behavior notes.

Developer Tools / Code Assistant 74 ↓ -1 89 days ago Details
#1947 ↓ -1
Meadow Mind

An open-source Python project that uses a local 7B diffusion language model to make real-time Gymnasium decisions with zero training. The README shows install steps, a code example, and benchmark-style results across several environments.

AI / Reinforcement Learning 74 ↓ -1 92 days ago Details

A shareable report for AI coding usage that summarizes how you work with tools like Claude Code, Codex, and Cursor, with sign-in via GitHub and an option to get printed copies.

Developer Tools / Code Assistant 74 ↓ -1 92 days ago Details
#1949 ↓ -1
LEVI

LEVI is a harness-first evolutionary framework for code and prompt optimization. It focuses on reducing LLM cost with diversity-preserving search, role-aware model routing, and a proxy benchmark, and presents comparative results against several existing systems.

Developer Tools / Code Assistant 74 ↓ -1 95 days ago Details
#1950 ↓ -1
Anthrosevka Mono

A custom Iosevka font build inspired by Anthropic Mono, with prebuilt releases and source build instructions.

Developer Tool / Font / Typeface Utility 74 ↓ -1 98 days ago Details

arXiv paper on distilling multi-agent debate into a single LLM with a two-stage fine-tuning pipeline. The abstract reports lower token use, comparable or better benchmark performance, and an analysis of agent-specific activation subspaces, with code linked from the page.

Research / AI/LLM Reasoning 74 ↓ -1 98 days ago Details

A blog post describing a small reinforcement-learning agent trained with PPO to play and beat a Pokelike/Pokerogue-style game, including the input representation, model architecture, and training approach.

Developer Tools / Code Assistant 74 ↓ -1 99 days ago Details

A detailed tutorial that explains what agent tools are and walks through a basic set of implementation patterns for an AI agent in Python, including bash, file, search, edit, and web fetch tools.

Writing / Copywriting 74 ↓ -1 101 days ago Details

A Codacy blog post explaining GitHub Copilot’s new AI Credits billing for code review and positioning Codacy AI Reviewer as an alternative with fixed per-seat pricing, static analysis, and pull-request review features.

Developer Tools / Code Assistant 74 ↓ -1 102 days ago Details

A tutorial showing how to build a release notes generator with the Cline SDK, including project structure, setup, custom tool design, and code walkthroughs.

Developer Tools / Code Assistant 74 ↓ -1 104 days ago Details
#1956 ↓ -1
x-commit

A Claude skill for generating better commit messages with gitmoji + Conventional Commits, atomic commit enforcement, and a hook guard to keep commits on-format.

Developer Tools / Git / Commit Automation 74 ↓ -1 106 days ago Details

A Superconductor blog post showing how background coding agents were used to reproduce, diagnose, and fix a Rails memory leak using derailed_benchmarks, with a reusable Agent Skill workflow included.

Developer Tools / Code Assistant 74 ↓ -1 106 days ago Details
#1958 ↓ -1
Ripgrep AI Policy

A short policy file for the ripgrep project that explains how contributors may use AI tools while contributing, and sets rules for issues, pull requests, and comments.

Writing / Copywriting 74 ↓ -1 106 days ago Details

A detailed article about the internal harness Nimbalyst built around Claude Code and Codex, covering context, provenance, capability, workflow, restraint, verification, visual interface, and coordination.

Developer Tools / Code Assistant 74 ↓ -1 108 days ago Details

A GitHub research project that measures how gpt-4.1 responds when asked to pick a random number between 1 and 100, using 10,000 API calls and comparing the results to a uniform baseline.

AI Research / Model Behavior Analysis 74 ↓ -1 109 days ago Details

A detailed write-up about using AI coding agents to build and tune a Rust multi-Paxos consensus engine, with concrete notes on contracts, spec-driven development, testing, and performance work.

AI Development / AI Coding Practices 74 ↓ -1 114 days ago Details
#1962 ↓ -1
Love Type Test

AI-driven relationship quiz that maps users into 16 and 32 love type profiles, with compatibility and self-reflection guidance.

Lifestyle / Quiz Tool / Relationship / Compatibility Quiz 74 ↓ -1 116 days ago Details

A minimal LLM agent packaged as a single HTML file. The README explains how to run it locally, set the LLM host, API key, and model, and use chat("hi") from the browser console.

Developer Tools / AI Agent Frameworks 74 ↓ -1 116 days ago Details
#1964 ↓ -1
claude-pee

A Rust CLI that wraps Claude Code for programmatic use, forwarding flags, injecting one-shot prompts, tailing session transcripts, and exiting automatically after the turn completes.

Developer Tools / CLI / Automation 74 ↓ -1 120 days ago Details
#1965 ↓ -1
3D-Agent

A Blender AI plugin roundup page that also spotlights 3D-Agent as a native in-Blender AI workflow tool. The page explains what the product does, how it fits into Blender workflows, and compares it with other related tools.

Productivity / Workflow Automation 74 ↓ -1 120 days ago Details
#1966 ↓ -1
showhn-rank

A Python pipeline that ranks 1,000 Show HN posts with an LLM judge and TrueSkill, then compares estimated merit against Hacker News points.

Developer Tools / AI Evaluation 74 ↓ -1 124 days ago Details

An open-source experiment that adds a small zero-initialized overlay layer to a frozen GPT-2 so its behavior can be adjusted at inference time without retraining the base model.

AI Developer Tool / Model Adaptation / Adapters 74 ↓ -1 128 days ago Details

A blog post about an experimental local LLM system where three qwen3.5:9b agents run autonomously, maintain stress states, self-modify capabilities, and coordinate through a shared OS-like layer.

Developer Tools / Code Assistant 74 ↑ +1 129 days ago Details