AgentDish directory
reasoning
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#872
↓ -6
LLM-Tests
An open-source benchmark for testing whether LLMs can follow long arithmetic derivations without using tools. It compares models on recall-proof inputs, logs digit accuracy, and reports results across multiple endpoints and model families. |
Developer Tools / Code Assistant | 84 | ↓ -6 | 30 days ago | Details |
|
#884
↓ -6
Monologue by Waterr
Research preview for a voice-agent reasoning harness that lets a realtime assistant think between turns without adding in-band latency. The page includes benchmark results, pricing, example behavior, and a description of the architecture behind Waterr’s AI meetings. |
Developer Tools / Code Assistant | 84 | ↓ -6 | 38 days ago | Details |
|
#1238
↓ -3
skills-for-humanity
A Claude Code skill pack with 171 structured reasoning methods drawn from historical thinkers, designed to route different problem types to specific procedures. |
Developer Tools / AI Development Tools | 83 | ↓ -3 | 108 days ago | Details |
|
Apple Machine Learning Research paper proposing LaDiR, a reasoning framework that combines a VAE-based latent space with latent diffusion to improve LLM text reasoning and iterative refinement. |
AI Research / LLM Reasoning | 78 | ↑ +5 | 128 days ago | Details |
|
#1912
↓ -1
AntiAgent
AntiAgent is a human-in-the-loop planning tool that uses AI to ask sharp questions, surface priorities, and turn messy thoughts into a step-by-step execution plan. |
AI Productivity / Decision Support | 74 | ↓ -1 | 7 days ago | Details |