AgentDish directory
Research AI Tools
Accepted listings in this category.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#47
↓ -2
Primus AI Researcher – Free
Primus is an autonomous AI researcher that hypothesizes, reads papers, writes code, runs experiments on compute, and drafts research papers. The page shows example tasks, published-paper claims, waitlist access, and positioning for ML research workflows. |
Research / Knowledge Work | 90 | ↓ -2 | 33 days ago | Details |
|
An interactive dashboard that analyzes New York Times coverage since 2000 using the NYT Archive API, with views for reporters, beats, sections, subjects, geography, obituaries, and corrections. |
Research / Data Visualization | 89 | ↑ +500 | 129 days ago | Details |
|
#340
↓ -3
CAD-Bench
An open benchmark and leaderboard for AI CAD agents, with 308 prompts across 20 categories and layered scoring for geometry, engineering, manufacturability, and cognition. |
Research / Knowledge Work | 88 | ↓ -3 | 126 days ago | Details |
|
A research article from Applied Compute on how agentic, tool-using workloads differ from traditional LLM benchmarks, with production observations, workload profiles, and an open-source harness for replaying traces. |
Research / Knowledge Work | 87 | ↓ -119 | 129 days ago | Details |
|
#508
↑ +2
ThoughtDAG
An open-source, local-first canvas for editing LLM context as a graph. It lets users branch, prune, merge, and inspect the exact context sent to a model, with a desktop app, web demo, and support for Ollama and OpenAI-compatible endpoints. |
Research / Knowledge Work | 86 | ↑ +2 | 28 days ago | Details |
|
#581
↑ +2
UnderstandDocs
UnderstandDocs is a document analysis tool that summarizes pasted text or uploaded files, flags risks, extracts important dates, and simplifies dense language. The page shows a working analysis form, supported file types, a privacy/no-storage claim, and an example output for a tenancy agreement. |
Research / Knowledge Work | 86 | ↑ +2 | 66 days ago | Details |
|
#584
↑ +2
A global workspace in language models
Anthropic research article on Claude interpretability, presenting evidence for a J-space and a global-workspace-like mechanism in language models. |
Research / Interpretability | 86 | ↑ +2 | 67 days ago | Details |
|
#596
↑ +2
ZUSE Automat Agent
A deterministic Python project for empirical law discovery in elementary cellular automata, with simulation, discovery, reproducibility guides, and published preprints. |
Research / Scientific Discovery | 86 | ↑ +2 | 75 days ago | Details |
|
#679
↑ +1007
Alignment Whack-a-Mole
A research code repository for studying how fine-tuning can trigger verbatim recall of copyrighted books in large language models. It includes preprocessing, fine-tuning, generation, and memorization-evaluation scripts, with setup notes and example data. |
Research / Copywriting | 86 | ↑ +1007 | 130 days ago | Details |
|
#850
↓ -6
Valovest
Valovest is an AI stock sentiment analysis tool that synthesizes analyst reports, earnings calls, news, and social sentiment into a short investment brief. The page shows timeframe filters, stock examples, a weekly featured stock, and outputs like overall sentiment, bull vs. bear arguments, and key themes. |
Research / Knowledge Work | 84 | ↓ -6 | 19 days ago | Details |
|
#960
↓ -6
Clusy
Clusy is an agent-native notebook platform for ML and data science that lets users describe a goal in plain language and have the system source data, set up experiments, run notebook cells, and return editable results in the cloud. |
Research / Knowledge Work | 84 | ↓ -6 | 65 days ago | Details |
|
arXiv paper describing QUEST, an open family of deep research agents from 2B to 35B parameters, plus a synthetic-task training recipe and released models, data, and scripts. |
Research / AI Agents | 83 | ↓ -3 | 109 days ago | Details |
|
#1241
↓ -3
wwwatch
A daily AI intelligence journal for builders, covering notable model, tooling, and release updates in a short sourced digest. |
Research / Knowledge Work | 83 | ↓ -3 | 113 days ago | Details |
|
#1245
↓ -3
Physics AI
Physics AI is a physics homework and study tool that solves problems from photos or typed prompts, with step-by-step explanations, tutor mode, and visual breakdowns for diagrams and vectors. |
Research / Knowledge Work | 83 | ↓ -3 | 114 days ago | Details |
|
#1266
↑ +194
Q2 2026 MCP Ecosystem Health
A research report on the current MCP ecosystem, with live crawl numbers, verification rates, category breakdowns, and examples of both strong and weak MCP-positive sites. |
Research / AI research | 83 | ↑ +194 | 129 days ago | Details |
|
#1313
↓ -2
Measuring Autonomous AI Research
A research post from Prime Intellect comparing 153 autonomous runs across 18 frontier models on a nanoGPT optimizer speedrun. It presents the setup, harness, results, and discussion around autonomous AI research performance. |
Research / AI Research Evaluation | 82 | ↓ -2 | 27 days ago | Details |
|
#1400
↓ -2
The Cascade Graph
An interactive, cited knowledge graph that maps physical, geopolitical, and economic constraints across the global economy, with node evidence pages, feedback loops, and an accessible text index. |
Research / Knowledge Work | 82 | ↓ -2 | 81 days ago | Details |
|
#1415
↓ -2
BigTech AI News
Chrome extension that tracks major AI companies, pulls in AI news and research, and generates daily summaries with Gemini, including language-aware summaries and article deep dives. |
Research / Knowledge Work | 82 | ↓ -2 | 94 days ago | Details |
|
#1455
↑ +228
ShadowBrokers
AI-powered trade signal product for retail traders that turns financial news into ranked trade plans with entries, stops, targets, and tracked accuracy. |
Research / Knowledge Work | 82 | ↑ +228 | 129 days ago | Details |
|
#1482
↑ +2
NanoGPT Speedrun Frontier
A research page comparing 153 autonomous runs across 18 frontier models on the nanoGPT optimizer speedrun, with rankings, trajectories, token/compute stats, and equal-budget comparisons. |
Research / Knowledge Work | 81 | ↑ +2 | 20 days ago | Details |
|
A research article from Plicara Labs analyzing 1.9 million GitHub agent skills and how often they include code, with breakdowns by language, mention-to-code gaps, and writing language. |
Research / Data Analysis | 80 | ↓ -1 | 13 days ago | Details |
|
#1539
↓ -1
Canvas Chat
Canvas Chat is a tree-based LLM chat interface for branching, comparing, and organizing conversations on an infinite canvas using your own API keys. |
Research / Knowledge Work | 80 | ↓ -1 | 16 days ago | Details |
|
A position paper arguing that AI alignment methods can be repurposed for censorship and manipulation, with examples across pre-training, post-training, and inference-time controls. |
Research / Knowledge Work | 79 | ↑ +2 | 66 days ago | Details |
|
A DeepMind technical report PDF about double-blind AI evaluations and confidentiality issues in AI safety research. |
Research / AI Safety | 78 | ↑ +6 | 10 days ago | Details |