AgentDish directory

AI Research AI Tools

Accepted listings in this category.

Listing Category Score Trend Checked
#576 ↑ +2
Prometheus

An autonomous research system that runs on a single workstation and aggressively checks its own claims with adversarial self-verification, replication, and calibration audits.

AI Research / Autonomous Research Systems 86 ↑ +2 64 days ago Details
#833 ↓ -6
Keenable SELECT

An MCP-based research agent that searches live web data through SQL and publishes reports with the full query and tool trajectory.

AI Research / Web Search / Data Extraction 84 ↓ -6 10 days ago Details
#1124 ↓ -3
Uno

Uno is a research repository for speeding up LLM inference with discrete diffusion and lossless multi-token decoding. The repo includes inference, training, and evaluation code, plus installation steps, checkpoints, and example workflows.

AI Research / LLM Efficiency / Inference 83 ↓ -3 8 days ago Details

An arXiv paper describing an open library of 163 procedural skills for research agents across 16 scientific areas, with versioned instruction files and accompanying reference material or runnable scripts.

AI Research / Research Agents 82 ↓ -2 9 days ago Details
#1398 ↓ -2
Socrates

Open-source multi-agent protocol for AI research agents. It pairs a tool-using Scientist with a question-only advisor that can only ask questions and approve plans, and the README includes quick-start setup plus notes on reproducing results on MLE-bench/Kaggle tasks.

AI Research / Multi-agent systems 82 ↓ -2 79 days ago Details

A research post from Telem AI measuring cold-start latency across nine web-search APIs, with cache behavior, tail latency, and agent-query comparisons.

AI Research / Benchmarking / Latency Research 81 ↑ +2 20 days ago Details
#1512 ↑ +2
EuroMesh

A sourced model and short report exploring whether Europe could train a sovereign frontier AI model using public compute it already owns, with reproducible code, datasets, and a PDF report.

AI Research / Analysis / Reports 81 ↑ +2 89 days ago Details
#1529 ↓ -64
MarCognity-AI

An open-source research framework for structured LLM evaluation, claim verification, and source-grounded reflective reasoning. The repo describes modular components for retrieval, semantic scoring, skeptical claim checking, and benchmark-style epistemic assessment.

AI Research / Evaluation / Verification Framework 81 ↓ -64 129 days ago Details

Google Research article describing Science One Framework, an autonomous research prototype focused on verifiable AI-generated research through evidence chains and claim verification.

AI Research / Autonomous Research 80 ↓ -1 34 days ago Details
#1627 ↑ +2
Two Eyes Arguing

An interactive AI experiment where two photorealistic eyes argue about everyday moral scenarios, grounded in the sociology of justification and symbolic boundaries.

AI Research / Interactive Demo 79 ↑ +2 43 days ago Details

A detailed technical write-up about training a 3.8B-parameter language model from scratch, including setup, throughput improvements, ablations, benchmarks, and cost/performance results.

AI Research / LLM Training / Write-up 78 ↑ +6 2 days ago Details
#1749 ↑ +6
MiroThinker

MiroThinker is a science-focused AI research app that emphasizes prediction, verification, and evidence-backed answers. The page also points to a MiroMind app and suggests use cases across finance, medicine, and regulation.

AI Research / Deep Research Agent 78 ↑ +6 92 days ago Details

arXiv paper describing AVA, a GenAI platform for policy and development research built on 4,000+ World Bank reports. The abstract highlights multilingual support, evidence-based synthesis, citation verifiability, and reasoned abstention when queries cannot be supported.

AI Research / Trustworthy Generative AI 78 ↑ +6 107 days ago Details

Agora-1 is a multi-agent world model from Odyssey that simulates shared real-time environments for up to four participants, human or AI, with a focus on gaming, robotics, reinforcement learning, and foundation model research.

AI Research / World Models 78 ↑ +6 116 days ago Details

Apple Machine Learning Research paper proposing LaDiR, a reasoning framework that combines a VAE-based latent space with latent diffusion to improve LLM text reasoning and iterative refinement.

AI Research / LLM Reasoning 78 ↑ +5 129 days ago Details
#1806 → 0
Maith

Maith is an open-source research workspace for exploring open mathematics with AI while keeping proof standards explicit and checkable.

AI Research / Mathematics 77 → 0 53 days ago Details
#1883 ↓ -1
Hyperagents

Research paper introducing hyperagents, a self-referential agent framework that combines a task agent and a meta agent into one editable program. The abstract describes a DGM-based system that improves both task performance and its own improvement process across domains.

AI Research / Self-Improving Agents 76 ↓ -1 112 days ago Details

A GitHub research project documenting a long-form, multi-model analysis of LLM behavior across Claude, Gemini, ChatGPT, and Grok. The repo includes an executive summary, screenplay, technical white paper, and archive of logs and chat records.

AI Research / LLM Evaluation & Analysis 75 → 0 109 days ago Details

A research page describing a verified circle-packing result produced by a frozen language model with an external memory loop, with published trace, attempts archive, and repository code/data.

AI Research / LLM Experiments 74 ↓ -1 12 days ago Details

A GitHub research project that measures how gpt-4.1 responds when asked to pick a random number between 1 and 100, using 10,000 API calls and comparing the results to a uniform baseline.

AI Research / Model Behavior Analysis 74 ↓ -1 110 days ago Details