AgentDish directory

performance

Accepted listings with this tag.

Listing Category Score Trend Checked

A web calculator for estimating whether local LLMs fit on specific hardware and how fast they may run. It covers VRAM, token throughput, latency, power, and cost across multiple GPUs and devices.

Developer Tools / Code Assistant 88 ↓ -3 25 hours ago Details
#179 ↓ -3
autotune

A drop-in proxy for Ollama that automatically tunes local LLM requests to reduce RAM use and speed up first-token latency. The page shows install steps, benchmark results, a live dashboard, and a local-only admin UI.

Developer Tools / Code Assistant 88 ↓ -3 37 days ago Details

LiteLLM is moving its AI gateway to Rust and claims major gains in throughput, memory use, and request overhead while keeping the same config, API, and provider coverage.

AI Gateway / Infrastructure 81 ↑ +2 45 days ago Details