AgentDish directory
AI research
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#93
↑ +258
prxhub
prxhub is an open registry for AI research bundles, built around verifiable .prx artifacts, search, provenance, and agent-friendly publishing via MCP and CLI. |
Developer Tools / Code Assistant | 90 | ↑ +258 | 129 days ago | Details |
|
#1124
↓ -3
Uno
Uno is a research repository for speeding up LLM inference with discrete diffusion and lossless multi-token decoding. The repo includes inference, training, and evaluation code, plus installation steps, checkpoints, and example workflows. |
AI Research / LLM Efficiency / Inference | 83 | ↓ -3 | 7 days ago | Details |
|
#1266
↑ +194
Q2 2026 MCP Ecosystem Health
A research report on the current MCP ecosystem, with live crawl numbers, verification rates, category breakdowns, and examples of both strong and weak MCP-positive sites. |
Research / AI research | 83 | ↑ +194 | 129 days ago | Details |
|
An arXiv paper describing an open library of 163 procedural skills for research agents across 16 scientific areas, with versioned instruction files and accompanying reference material or runnable scripts. |
AI Research / Research Agents | 82 | ↓ -2 | 8 days ago | Details |
|
A research post from Telem AI measuring cold-start latency across nine web-search APIs, with cache behavior, tail latency, and agent-query comparisons. |
AI Research / Benchmarking / Latency Research | 81 | ↑ +2 | 19 days ago | Details |
|
#1504
↑ +2
AI Agent Safety & Alignment — Agent Bayes
An early-access AI research assistant that presents a cited mindmap of AI agent safety and alignment research, with expandable nodes, source-backed claims, and built-in verification checks. |
Developer Tools / Code Assistant | 81 | ↑ +2 | 72 days ago | Details |
|
#1512
↑ +2
EuroMesh
A sourced model and short report exploring whether Europe could train a sovereign frontier AI model using public compute it already owns, with reproducible code, datasets, and a PDF report. |
AI Research / Analysis / Reports | 81 | ↑ +2 | 88 days ago | Details |
|
#1529
↓ -64
MarCognity-AI
An open-source research framework for structured LLM evaluation, claim verification, and source-grounded reflective reasoning. The repo describes modular components for retrieval, semantic scoring, skeptical claim checking, and benchmark-style epistemic assessment. |
AI Research / Evaluation / Verification Framework | 81 | ↓ -64 | 129 days ago | Details |
|
Google Research article describing Science One Framework, an autonomous research prototype focused on verifiable AI-generated research through evidence chains and claim verification. |
AI Research / Autonomous Research | 80 | ↓ -1 | 33 days ago | Details |
|
#1627
↑ +2
Two Eyes Arguing
An interactive AI experiment where two photorealistic eyes argue about everyday moral scenarios, grounded in the sociology of justification and symbolic boundaries. |
AI Research / Interactive Demo | 79 | ↑ +2 | 42 days ago | Details |
|
#1683
↑ +6
Training a 3.8B LLM to 0.384 CORE for $998
A detailed technical write-up about training a 3.8B-parameter language model from scratch, including setup, throughput improvements, ablations, benchmarks, and cost/performance results. |
AI Research / LLM Training / Write-up | 78 | ↑ +6 | 43 hours ago | Details |
|
#1749
↑ +6
MiroThinker
MiroThinker is a science-focused AI research app that emphasizes prediction, verification, and evidence-backed answers. The page also points to a MiroMind app and suggests use cases across finance, medicine, and regulation. |
AI Research / Deep Research Agent | 78 | ↑ +6 | 91 days ago | Details |
|
#1766
↑ +6
Multi-Agent is a snake oil
An opinionated write-up on where multi-agent systems have and have not delivered value, with concrete comparisons across coding, images, CAD, and BIM, plus a description of Blade’s evidence-first architecture. |
Writing / Copywriting | 78 | ↑ +6 | 108 days ago | Details |
|
#1773
↑ +6
Agora-1: The Multi-Agent World Model
Agora-1 is a multi-agent world model from Odyssey that simulates shared real-time environments for up to four participants, human or AI, with a focus on gaming, robotics, reinforcement learning, and foundation model research. |
AI Research / World Models | 78 | ↑ +6 | 115 days ago | Details |
|
Apple Machine Learning Research paper proposing LaDiR, a reasoning framework that combines a VAE-based latent space with latent diffusion to improve LLM text reasoning and iterative refinement. |
AI Research / LLM Reasoning | 78 | ↑ +5 | 128 days ago | Details |
|
#1806
→ 0
Maith
Maith is an open-source research workspace for exploring open mathematics with AI while keeping proof standards explicit and checkable. |
AI Research / Mathematics | 77 | → 0 | 52 days ago | Details |
|
#1883
↓ -1
Hyperagents
Research paper introducing hyperagents, a self-referential agent framework that combines a task agent and a meta agent into one editable program. The abstract describes a DGM-based system that improves both task performance and its own improvement process across domains. |
AI Research / Self-Improving Agents | 76 | ↓ -1 | 111 days ago | Details |
|
#1916
↓ -1
Extrapolation Under an Exact Verifier
A research page describing a verified circle-packing result produced by a frozen language model with an external memory loop, with published trace, attempts archive, and repository code/data. |
AI Research / LLM Experiments | 74 | ↓ -1 | 11 days ago | Details |
|
#1960
↓ -1
GPT Guesses Between 1 and 100
A GitHub research project that measures how gpt-4.1 responds when asked to pick a random number between 1 and 100, using 10,000 API calls and comparing the results to a uniform baseline. |
AI Research / Model Behavior Analysis | 74 | ↓ -1 | 109 days ago | Details |