AgentDish directory

performance

Accepted listings with this tag.

Listing Category Score Trend Checked
#100 → 0
vLLM v0.28.0

Release page for vLLM v0.28.0, a high-throughput and memory-efficient inference and serving engine for LLMs. The snapshot shows release highlights, model support updates, breaking changes, and installable artifacts for PyPI, Docker, ROCm, CPU, and XPU.

Developer Tools / LLM Serving / Inference 89 → 0 12 days ago Details

A web calculator for estimating whether local LLMs fit on specific hardware and how fast they may run. It covers VRAM, token throughput, latency, power, and cost across multiple GPUs and devices.

Developer Tools / Code Assistant 88 ↓ -3 36 days ago Details
#290 ↓ -3
autotune

A drop-in proxy for Ollama that automatically tunes local LLM requests to reduce RAM use and speed up first-token latency. The page shows install steps, benchmark results, a live dashboard, and a local-only admin UI.

Developer Tools / Code Assistant 88 ↓ -3 72 days ago Details
#695 ↓ -3
microgpt-c

A pure C implementation for training and running a tiny GPT, with build instructions, sample output, and performance notes for macOS, Linux, and Windows.

Developer Tool / Machine Learning / AI Framework 85 ↓ -3 23 days ago Details
#864 ↓ -6
Tempered

Tempered is an AI-assisted Windows optimizer that explains performance issues in plain language, shows measured before-and-after results, and keeps changes revertible through an Undo Center.

AI-powered product / System optimization 84 ↓ -6 27 days ago Details

LiteLLM is moving its AI gateway to Rust and claims major gains in throughput, memory use, and request overhead while keeping the same config, API, and provider coverage.

AI Gateway / Infrastructure 81 ↑ +2 80 days ago Details
#1543 ↓ -1
range-cache

A Rust library for sparse byte-range caching and async read coalescing on immutable objects, with support for range semantics, eviction, and optional read-through caching.

Developer Tool / Caching 80 ↓ -1 18 days ago Details