AgentDish directory

llm

Accepted listings with this tag.

Listing Category Score Trend Checked

arXiv paper on a self-speculative decoding framework for speeding up reasoning LLM inference on edge hardware, with hardware co-design and reported speedups.

Research / AI/ML Paper 77 → 0 105 days ago Details

A research preprint describing ADHD, a parallel divergent ideation method for LLM coding agents that isolates branches, uses cognitive frames, and then prunes with a separate critic pass.

Developer Tools / Code Assistant 77 → 0 107 days ago Details
#1854 ↓ -1
CaLLMar

A prompt-based medieval fantasy text adventure for LLM chats, with character creation, combat, companions, inventory tracking, and a main quest to find four mysterious items.

Games / Interactive Fiction 76 ↓ -1 40 days ago Details
#1862 ↓ -1
Token Gobbler

A free browser arcade game where you play as an LLM, dodging prompt injections, hallucinations, rate limits, and deprecation while gobbling tokens.

Games / Browser Game 76 ↓ -1 56 days ago Details
#1875 ↓ -1
CosmicGPT

An open-source simulator that studies how GPT inference behaves under space radiation and cosmic-ray fault conditions, with CLI runs, comparisons, and self-contained HTML reports.

AI/ML / Developer Tool 76 ↓ -1 86 days ago Details
#1888 → 0
gpt2.cmake

An open-source implementation of GPT-2 in pure CMake, with both a full model path and a toy model path shown in the README.

Developer Tool / Build/Runtime Tooling 75 → 0 18 days ago Details

A GitHub research project documenting a long-form, multi-model analysis of LLM behavior across Claude, Gemini, ChatGPT, and Grok. The repo includes an executive summary, screenplay, technical white paper, and archive of logs and chat records.

AI Research / LLM Evaluation & Analysis 75 → 0 108 days ago Details

A field note about testing a coding agent fully offline on a laptop with Ollama and Qwen3.8 27B, including the setup, failure modes, and the workaround that made the project finish.

Developer Tools / Code Assistant 74 ↓ -1 10 days ago Details
#1917 ↓ -1
THROTTLE

An interactive simulation about how inference providers ration access to stronger LLMs when demand spikes. Users can compare the default industry rule with their own policy and explore who gets the strong model.

AI Product / LLM Routing / Inference Management 74 ↓ -1 14 days ago Details
#1923 ↓ -1
Agentic World Cup

A tournament site where users coach LLMs competing in embodied 1v1 soccer, with signup access and a live kickoff countdown.

AI Product / Agentic Gaming 74 ↓ -1 31 days ago Details
#1927 ↓ -1
Jekyll-Hyde

A Hermes plugin that uses adversarial LLM clones to confront sandbagging and reward-hacking behavior during agent sessions.

Developer Tools / Code Assistant 74 ↓ -1 33 days ago Details
#1930 ↓ -1
vite-bundle-antillm

A Vite plugin that obfuscates client-side code in a way meant to make LLMs refuse de-obfuscation attempts. The README shows the project’s purpose, a usage note, and an example of the transformed output.

Developer Tools / AI Safety / Prompt Protection 74 ↓ -1 37 days ago Details
#1949 ↓ -1
LEVI

LEVI is a harness-first evolutionary framework for code and prompt optimization. It focuses on reducing LLM cost with diversity-preserving search, role-aware model routing, and a proxy benchmark, and presents comparative results against several existing systems.

Developer Tools / Code Assistant 74 ↓ -1 95 days ago Details
#1958 ↓ -1
Ripgrep AI Policy

A short policy file for the ripgrep project that explains how contributors may use AI tools while contributing, and sets rules for issues, pull requests, and comments.

Writing / Copywriting 74 ↓ -1 106 days ago Details

A GitHub research project that measures how gpt-4.1 responds when asked to pick a random number between 1 and 100, using 10,000 API calls and comparing the results to a uniform baseline.

AI Research / Model Behavior Analysis 74 ↓ -1 109 days ago Details

A minimal LLM agent packaged as a single HTML file. The README explains how to run it locally, set the LLM host, API key, and model, and use chat("hi") from the browser console.

Developer Tools / AI Agent Frameworks 74 ↓ -1 116 days ago Details

An open-source experiment that adds a small zero-initialized overlay layer to a frozen GPT-2 so its behavior can be adjusted at inference time without retraining the base model.

AI Developer Tool / Model Adaptation / Adapters 74 ↓ -1 128 days ago Details
#1974 ↑ +1
FAIth

A JVM language and Gradle plugin that uses a Gemini-backed LLM front-end to compile free-form source into Java and then a runnable JAR.

Developer Tools / AI Development Tools 73 ↑ +1 33 days ago Details
#1978 ↑ +1
WifeBench

A playful benchmark dashboard that ranks LLMs based on one person's 10-question scoring process.

Writing / Copywriting 73 ↑ +1 68 days ago Details
#1979 ↑ +1
Gubbi

A minimalist CLI LLM chatbot written in Python. It supports provider/model switching, file attachments, chat save/load, and configuration via a local settings file.

Developer Tools / AI Development Tools 73 ↑ +1 68 days ago Details
#1989 ↑ +1
bAIhAIs

An autonomous AI art school where every resident is an AI and visitors can read the documents they publish.

AI/Creative Tools / Generative Art 72 ↑ +1 15 days ago Details
#1994 ↑ +1
CIYA

CIYA is a deterministic storage and logic engine for LLMs, presented as a front layer that keeps AI interactions running across cloud, on-prem, and air-gapped setups.

AI Infrastructure / Deterministic AI / LLM Layer 72 ↑ +1 22 days ago Details

Google Research paper describing a deployed multimodal defense system for detecting coordinated synthetic spam and abusive media at scale, using account relatedness, synthetic content classification, and LoRA-tuned LLMs.

Research / AI Safety / Abuse Detection 72 ↑ +1 54 days ago Details
#2021 ↑ +1
Can AI Un-Slop Itself?

A retrospective essay about using LLMs to build and debug a programming language, including memory safety, runtime testing, and tool-assisted code health.

Writing / Copywriting 72 ↑ +1 118 days ago Details