AgentDish directory

reasoning

Accepted listings with this tag.

Listing Category Score Trend Checked
#872 ↓ -6
LLM-Tests

An open-source benchmark for testing whether LLMs can follow long arithmetic derivations without using tools. It compares models on recall-proof inputs, logs digit accuracy, and reports results across multiple endpoints and model families.

Developer Tools / Code Assistant 84 ↓ -6 30 days ago Details
#884 ↓ -6
Monologue by Waterr

Research preview for a voice-agent reasoning harness that lets a realtime assistant think between turns without adding in-band latency. The page includes benchmark results, pricing, example behavior, and a description of the architecture behind Waterr’s AI meetings.

Developer Tools / Code Assistant 84 ↓ -6 38 days ago Details
#1238 ↓ -3
skills-for-humanity

A Claude Code skill pack with 171 structured reasoning methods drawn from historical thinkers, designed to route different problem types to specific procedures.

Developer Tools / AI Development Tools 83 ↓ -3 108 days ago Details

Apple Machine Learning Research paper proposing LaDiR, a reasoning framework that combines a VAE-based latent space with latent diffusion to improve LLM text reasoning and iterative refinement.

AI Research / LLM Reasoning 78 ↑ +5 128 days ago Details
#1912 ↓ -1
AntiAgent

AntiAgent is a human-in-the-loop planning tool that uses AI to ask sharp questions, surface priorities, and turn messy thoughts into a step-by-step execution plan.

AI Productivity / Decision Support 74 ↓ -1 7 days ago Details