AgentDish directory

Research AI Tools

Accepted listings in this category.

Listing Category Score Trend Checked

Primus is an autonomous AI researcher that hypothesizes, reads papers, writes code, runs experiments on compute, and drafts research papers. The page shows example tasks, published-paper claims, waitlist access, and positioning for ML research workflows.

Research / Knowledge Work 90 ↓ -2 33 days ago Details

An interactive dashboard that analyzes New York Times coverage since 2000 using the NYT Archive API, with views for reporters, beats, sections, subjects, geography, obituaries, and corrections.

Research / Data Visualization 89 ↑ +500 129 days ago Details
#340 ↓ -3
CAD-Bench

An open benchmark and leaderboard for AI CAD agents, with 308 prompts across 20 categories and layered scoring for geometry, engineering, manufacturability, and cognition.

Research / Knowledge Work 88 ↓ -3 126 days ago Details

A research article from Applied Compute on how agentic, tool-using workloads differ from traditional LLM benchmarks, with production observations, workload profiles, and an open-source harness for replaying traces.

Research / Knowledge Work 87 ↓ -119 129 days ago Details
#508 ↑ +2
ThoughtDAG

An open-source, local-first canvas for editing LLM context as a graph. It lets users branch, prune, merge, and inspect the exact context sent to a model, with a desktop app, web demo, and support for Ollama and OpenAI-compatible endpoints.

Research / Knowledge Work 86 ↑ +2 28 days ago Details
#581 ↑ +2
UnderstandDocs

UnderstandDocs is a document analysis tool that summarizes pasted text or uploaded files, flags risks, extracts important dates, and simplifies dense language. The page shows a working analysis form, supported file types, a privacy/no-storage claim, and an example output for a tenancy agreement.

Research / Knowledge Work 86 ↑ +2 66 days ago Details

Anthropic research article on Claude interpretability, presenting evidence for a J-space and a global-workspace-like mechanism in language models.

Research / Interpretability 86 ↑ +2 67 days ago Details
#596 ↑ +2
ZUSE Automat Agent

A deterministic Python project for empirical law discovery in elementary cellular automata, with simulation, discovery, reproducibility guides, and published preprints.

Research / Scientific Discovery 86 ↑ +2 75 days ago Details
#679 ↑ +1007
Alignment Whack-a-Mole

A research code repository for studying how fine-tuning can trigger verbatim recall of copyrighted books in large language models. It includes preprocessing, fine-tuning, generation, and memorization-evaluation scripts, with setup notes and example data.

Research / Copywriting 86 ↑ +1007 130 days ago Details
#850 ↓ -6
Valovest

Valovest is an AI stock sentiment analysis tool that synthesizes analyst reports, earnings calls, news, and social sentiment into a short investment brief. The page shows timeframe filters, stock examples, a weekly featured stock, and outputs like overall sentiment, bull vs. bear arguments, and key themes.

Research / Knowledge Work 84 ↓ -6 19 days ago Details
#960 ↓ -6
Clusy

Clusy is an agent-native notebook platform for ML and data science that lets users describe a goal in plain language and have the system source data, set up experiments, run notebook cells, and return editable results in the cloud.

Research / Knowledge Work 84 ↓ -6 65 days ago Details

arXiv paper describing QUEST, an open family of deep research agents from 2B to 35B parameters, plus a synthetic-task training recipe and released models, data, and scripts.

Research / AI Agents 83 ↓ -3 109 days ago Details
#1241 ↓ -3
wwwatch

A daily AI intelligence journal for builders, covering notable model, tooling, and release updates in a short sourced digest.

Research / Knowledge Work 83 ↓ -3 113 days ago Details
#1245 ↓ -3
Physics AI

Physics AI is a physics homework and study tool that solves problems from photos or typed prompts, with step-by-step explanations, tutor mode, and visual breakdowns for diagrams and vectors.

Research / Knowledge Work 83 ↓ -3 114 days ago Details

A research report on the current MCP ecosystem, with live crawl numbers, verification rates, category breakdowns, and examples of both strong and weak MCP-positive sites.

Research / AI research 83 ↑ +194 129 days ago Details

A research post from Prime Intellect comparing 153 autonomous runs across 18 frontier models on a nanoGPT optimizer speedrun. It presents the setup, harness, results, and discussion around autonomous AI research performance.

Research / AI Research Evaluation 82 ↓ -2 27 days ago Details
#1400 ↓ -2
The Cascade Graph

An interactive, cited knowledge graph that maps physical, geopolitical, and economic constraints across the global economy, with node evidence pages, feedback loops, and an accessible text index.

Research / Knowledge Work 82 ↓ -2 81 days ago Details
#1415 ↓ -2
BigTech AI News

Chrome extension that tracks major AI companies, pulls in AI news and research, and generates daily summaries with Gemini, including language-aware summaries and article deep dives.

Research / Knowledge Work 82 ↓ -2 94 days ago Details
#1455 ↑ +228
ShadowBrokers

AI-powered trade signal product for retail traders that turns financial news into ranked trade plans with entries, stops, targets, and tracked accuracy.

Research / Knowledge Work 82 ↑ +228 129 days ago Details

A research page comparing 153 autonomous runs across 18 frontier models on the nanoGPT optimizer speedrun, with rankings, trajectories, token/compute stats, and equal-budget comparisons.

Research / Knowledge Work 81 ↑ +2 20 days ago Details

A research article from Plicara Labs analyzing 1.9 million GitHub agent skills and how often they include code, with breakdowns by language, mention-to-code gaps, and writing language.

Research / Data Analysis 80 ↓ -1 13 days ago Details
#1539 ↓ -1
Canvas Chat

Canvas Chat is a tree-based LLM chat interface for branching, comparing, and organizing conversations on an infinite canvas using your own API keys.

Research / Knowledge Work 80 ↓ -1 16 days ago Details

A position paper arguing that AI alignment methods can be repurposed for censorship and manipulation, with examples across pre-training, post-training, and inference-time controls.

Research / Knowledge Work 79 ↑ +2 66 days ago Details

A DeepMind technical report PDF about double-blind AI evaluations and confidentiality issues in AI safety research.

Research / AI Safety 78 ↑ +6 10 days ago Details