AgentDish directory
llm
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#2
↑ +4
Claude API
Anthropic’s Claude API platform for building AI products, apps, and agent workflows with Claude models, built-in tools, pricing tiers, and developer docs. |
Development / API | 93 | ↑ +4 | 118 days ago | Details |
|
#7
↑ +298
WebLLM
WebLLM is a high-performance in-browser LLM inference engine that runs locally in the browser with WebGPU acceleration. It exposes an OpenAI-compatible API, supports streaming and JSON mode, and includes examples for building chat apps and browser extensions. |
Developer Tool / AI SDK / In-browser LLM inference | 92 | ↑ +298 | 118 days ago | Details |
|
#11
↓ -3
gemma4.c
A pure C runtime for Gemma 4 E2B CPU inference, with export, benchmarking, and numerical validation tools included. |
Developer Tools / Machine Learning / Inference Runtime | 91 | ↓ -3 | 3 days ago | Details |
|
#38
↓ -2
prompt-scrub
A local-first Node.js utility that redacts emails, phone numbers, paths, secrets, URLs, and other identifiers from LLM prompts, then rehydrates responses locally using session-based mappings. |
AI Developer Tool / Prompt privacy / PII redaction | 90 | ↓ -2 | 31 days ago | Details |
|
#41
↓ -2
TokenTown
An interactive browser city that visualizes how an LLM works, with live transformer mechanics such as tokenization, attention, KV cache, residual stream, and sampling. |
AI Education / Interactive Explainer | 90 | ↓ -2 | 34 days ago | Details |
|
#49
↓ -2
WakaWiki
WakaWiki is a Rust CLI that generates and keeps agent-facing documentation up to date for a codebase. It supports multiple model providers, a zero-LLM scan mode, incremental updates, and AGENTS.md integration, with CI-friendly GitHub Actions usage documented in the README. |
Developer Tools / AI Coding | 90 | ↓ -2 | 58 days ago | Details |
|
#73
↓ -1
Docker AI Stack
A self-hosted Docker Compose stack for running local AI services including Ollama, LiteLLM, Whisper, Kokoro, embeddings, and an MCP Gateway. |
Developer Tools / Self-hosting / AI Infrastructure | 90 | ↓ -1 | 118 days ago | Details |
|
#75
↑ +338
Open Bias
Open Bias is an open-source reliability harness for LLM apps and agents that enforces rules at runtime. It sits between your app and an LLM provider, using RULES.md policies to trace, block, or fix off-policy behavior. |
Developer Tools / Code Assistant | 90 | ↑ +338 | 118 days ago | Details |
|
#80
→ 0
vLLM v0.28.0
Release page for vLLM v0.28.0, a high-throughput and memory-efficient inference and serving engine for LLMs. The snapshot shows release highlights, model support updates, breaking changes, and installable artifacts for PyPI, Docker, ROCm, CPU, and XPU. |
Developer Tools / LLM Serving / Inference | 89 | → 0 | 2 days ago | Details |
|
#94
→ 0
Shoehorn
Shoehorn is a Rust library and CLI for fitting GGUF LLMs into available Mac VRAM by choosing a per-tensor mixed-precision quantization plan. It can probe usable GPU memory, account for KV cache and compute buffers, generate or use an imatrix, write a fitted GGUF, and launch llama.cpp inference. |
Developer Tools / AI/ML Developer Tools | 89 | → 0 | 18 days ago | Details |
|
#100
→ 0
Guided Review
Chrome extension that helps humans review AI-generated GitHub PRs by clustering diffs into review units and adding short AI summaries. |
Developer Tools / Code Assistant | 89 | → 0 | 26 days ago | Details |
|
#108
→ 0
esp32-ai
Open-source project showing how to run a 28.9M-parameter language model on an ESP32-S3 microcontroller, with on-device inference, firmware, wiring, training code, and measured results. |
Developer Tools / AI / Edge ML | 89 | → 0 | 37 days ago | Details |
|
#114
→ 0
Telemetry
Telemetry is an observability platform for AI and LLM apps that tracks model calls, tool steps, retrievals, tokens, cost, latency, and errors. It supports OpenTelemetry ingest, TypeScript and Python SDKs, and integrations like Vercel AI SDK and LangChain. |
Developer Tools / Code Assistant | 89 | → 0 | 41 days ago | Details |
|
#115
→ 0
Headroom
Headroom is an open-source context compression layer for AI agents. It compresses tool outputs, logs, files, RAG chunks, and conversation history before they reach the LLM, with support for a library, proxy, and MCP server. |
Developer Tools / AI/LLM Infrastructure | 89 | → 0 | 41 days ago | Details |
|
#142
→ 0
Redact
Browser extension that scans pastes for credentials and PII before they reach LLM chat sites, with local on-device inference and no network calls. |
Developer Tool / Browser Extension | 89 | → 0 | 90 days ago | Details |
|
#158
↓ -79
AutoRound
AutoRound is an open-source quantization toolkit for LLMs and VLMs, focused on high-accuracy low-bit inference across CPU, XPU, CUDA, and multiple deployment backends. |
Developer Tools / AI Infrastructure | 89 | ↓ -79 | 119 days ago | Details |
|
A web calculator for estimating whether local LLMs fit on specific hardware and how fast they may run. It covers VRAM, token throughput, latency, power, and cost across multiple GPUs and devices. |
Developer Tools / Code Assistant | 88 | ↓ -3 | 26 days ago | Details |
|
#210
↓ -3
Grafana AI SDK
Go SDK for building AI-backed apps and endpoints with streaming, tool calling, structured output, and React frontend compatibility. |
Developer Tools / AI SDK | 88 | ↓ -3 | 33 days ago | Details |
|
#215
↓ -3
arcade-js
A repository that uses AI agents to translate arcade game machine code into idiomatic JavaScript and validate the result pixel-for-pixel against MAME. The page shows real project status, test methodology, and examples of completed and in-progress ports. |
AI Development / AI-powered code generation | 88 | ↓ -3 | 36 days ago | Details |
|
#236
↓ -3
NeatContext
NeatContext is a local-first desktop workspace that adds team-specific knowledge and read-only internal tools to LLMs so they can give grounded operational answers. The page shows incident-response examples, Markdown-based profiles, MCP integration, and support for multiple model providers. |
AI Productivity / Developer Workflow | 88 | ↓ -3 | 50 days ago | Details |
|
#239
↓ -3
LocalMask
LocalMask is a local-first privacy layer for AI coding that scans repos on-device, masks secrets and PII into reversible tokens, and forwards only safe content to models like Claude, GPT, or Gemini. It also supports masked repo mirrors, an AI proxy, and a CLI/MCP workflow. |
Developer Tools / Privacy / Security | 88 | ↓ -3 | 51 days ago | Details |
|
#258
↓ -3
promptctl
CLI tool for versioning, diffing, rolling back, and searching AI prompts, with Git-style workflows for prompt management. |
Developer Tools / CLI | 88 | ↓ -3 | 68 days ago | Details |
|
#263
↓ -3
Latitude
Open-source AI agent monitoring platform with tracing, issue grouping, evals, semantic search, alerts, and MCP access for coding agents. |
Developer Tools / AI Observability | 88 | ↓ -3 | 70 days ago | Details |
|
#268
↓ -3
BeamWeaver
An OTP-native Elixir library for building AI agents, durable LLM workflows, retrieval pipelines, and production LLM services. |
Developer Tools / AI Agent Framework | 88 | ↓ -3 | 74 days ago | Details |