AgentDish directory

llm

Accepted listings with this tag.

Listing Category Score Trend Checked
#1638 ↑ +2
Lagotto Meter

Lagotto Meter is a beta web tool that checks how an AI agent perceives a website versus what the site claims about itself. It positions itself as a semantic truth/readability analyzer with a reproducible analysis method and public tools.

Developer Tool / AI Evaluation 79 ↑ +2 52 days ago Details

A pull request adding AMD GPU support to tiny-vLLM through ROCm/HIP while keeping the existing CUDA build path unchanged. The snapshot describes the compatibility header, CMake option, architecture selection, and validation on multiple AMD GPUs.

Developer Tool / LLM inference engine 79 ↑ +2 74 days ago Details
#1675 ↑ +2
Laze

An experimental programming language and compiler designed for LLM-written code, with a Python-based toolchain that compiles to native macOS binaries. The repo also includes a full NES emulator written in Laze and usage examples for both the compiler and emulator.

Developer Tool / Language Tooling 79 ↑ +2 122 days ago Details
#1677 ↑ +2
GenZ LLM

A small post-trained Qwen2.5-0.5B-Instruct model tuned to write in Gen Z slang, with training and inference notebooks plus dataset files in the repo.

AI Models / Fine-Tuned LLMs 79 ↑ +2 123 days ago Details
#1679 ↑ +2
ModelDocker

Windows desktop chat client for OpenRouter with streaming completions, model browsing, multi-session history, local storage, and PySide6 UI.

Developer Tools / Desktop App 79 ↑ +2 124 days ago Details
#1680 ↑ +2
context-editor-agent

An open-source desktop client for editing LLM context with a separate AI model, context maps, compression tools, and revision rollback.

Developer Tools / AI Development 79 ↑ +2 126 days ago Details

An Apple Silicon–optimized inference build of Bonsai 1.7B with custom Metal kernels, benchmark results, quick-start instructions, and a bundled OpenAI-compatible server.

Developer Tools / Code Assistant 79 ↓ -223 128 days ago Details

A detailed technical write-up about training a 3.8B-parameter language model from scratch, including setup, throughput improvements, ablations, benchmarks, and cost/performance results.

AI Research / LLM Training / Write-up 78 ↑ +6 45 hours ago Details
#1685 ↑ +6
Quick Markdown (QMD)

A lightweight Markdown variant designed to reduce token usage for LLM consumption using single-character markers and context-aware parsing.

Developer Tools / Prompt Engineering 78 ↑ +6 4 days ago Details

A DeepSeek Harness plugin that uses an LLM to verify and rank candidate outputs with select, compare, track, and rollout tools.

Developer Tools / AI DevTools 78 ↑ +6 22 days ago Details
#1702 ↑ +6
Vomit

Vomit is a Go-based CLI that converts Claude output into readable English by sending it through a separate local LLM. It supports install/setup commands, a scrub mode for Claude hooks, and sidecar commands for listing and tailing sessions.

Developer Tools / AI Developer Tool 78 ↑ +6 22 days ago Details

A pricing analysis article comparing batch API discounts across Google, OpenAI, Anthropic, xAI, and OpenRouter, with a focus on where batch rates differ from standard pricing.

Writing / Copywriting 78 ↑ +6 28 days ago Details
#1705 ↑ +6
HyperSAE

HyperSAE is a Python package for mechanistic interpretability that trains hyperbolic sparse autoencoders on LLM activations. The page includes installation steps, a quickstart example, benchmark tables, and a short software architecture overview.

Developer Tools / AI/ML Frameworks 78 ↑ +6 30 days ago Details
#1709 ↑ +6
OneRingAI

OneRingAI is an open-source TypeScript library for building AI agents with integrations and graph memory. The repository shows a substantial codebase with docs, setup files, testing, MCP-related material, and references to supported models and multimodal capabilities.

Developer Tools / AI Agent Framework 78 ↑ +6 33 days ago Details

An open-source project for running a small LLM and limited agent workflows on ESP32-P4 hardware, with model details, architecture notes, examples, and setup instructions.

AI Developer Tool / Edge AI / Embedded LLM 78 ↑ +6 36 days ago Details
#1735 ↑ +6
LLM Tools

A Python tool manager for AI agents that uses an LLM-tools.txt registry to install, pin, describe, and execute tools with one predictable contract.

Developer Tools / AI Tooling 78 ↑ +6 73 days ago Details
#1736 ↑ +6
EdgeSync-LLM

A GitHub repository for an engine-agnostic KV cache fragment system for on-device LLM inference, with Android and Go components, adapters for llama.cpp/MLC-LLM/ONNX Runtime, and benchmark and monitoring code.

Developer Tools / AI/LLM Inference 78 ↑ +6 73 days ago Details

A local-first desktop environment for dispatching coding tasks to any model and any agent, with diff review, terminal streaming, task history, and per-project configuration.

Developer Tools / AI Desktop Apps 78 ↑ +6 94 days ago Details
#1772 ↑ +6
id-agent

An open-source ID library that generates human-readable, token-efficient IDs for AI agents and LLM workflows, with parsing, validation, deterministic IDs, and alias mapping utilities.

Developer Tools / AI Developer Tools 78 ↑ +6 115 days ago Details
#1774 ↑ +6
MaragingLoop

An experimental autonomous bare-metal OS agent that uses an LLM, VirtualBox, and screenshot inspection to iterate on low-level code, boot a VM, and refine behavior from execution results.

Developer Tools / AI Agent Framework 78 ↑ +6 117 days ago Details
#1777 ↑ +6
markdown-parser

A TypeScript markdown parser built for incremental streaming, aimed at parsing LLM markdown output on the server or client. It exposes a typed AST, supports CommonMark and GFM tables, and includes examples for full and streaming parsing.

Developer Tools / Parsing 78 ↑ +6 119 days ago Details

Apple Machine Learning Research paper proposing LaDiR, a reasoning framework that combines a VAE-based latent space with latent diffusion to improve LLM text reasoning and iterative refinement.

AI Research / LLM Reasoning 78 ↑ +5 128 days ago Details

Public evaluation code for an Agent Memory Leaderboard, including answer-generation and scoring contracts for comparing LLM agent memory systems.

Developer Tool / AI Evaluation / Benchmarking 77 → 0 40 days ago Details

A write-up of a custom French-learning system built around spaced repetition, grammar tracking, and a voice practice app. It explains the problem, the workflow, and the model/API stack used to keep costs low.

Productivity / Workflow Automation 77 → 0 81 days ago Details