AgentDish directory
LLM infrastructure
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#324
↓ -3
Mnemo
Local-first AI memory layer for LLM apps with persistent knowledge graph storage, entity extraction, semantic retrieval, and API endpoints for ingesting and retrieving context. |
Developer Tools / Code Assistant | 88 | ↓ -3 | 99 days ago | Details |
|
#661
↑ +2
Adola
Adola offers Rose 1, a prompt compression API that trims noisy context before LLM calls while aiming to preserve answer quality. The page shows quickstart code, benchmark claims, and use cases for agent traces, RAG retrieval, prompt gateways, and support copilots. |
Developer Tools / AI API | 86 | ↑ +2 | 122 days ago | Details |
|
#701
↓ -3
kvcachescope
A logical KV cache profiler and leak detector for PagedAttention engines like vLLM and SGLang, with a dashboard, live engine hook, CI checks, and Perfetto trace export. |
Developer Tools / Code Assistant | 85 | ↓ -3 | 27 days ago | Details |
|
#1476
↑ +2
TokenGO
TokenGO is a self-operated inference API for open-weight LLMs and video models. It offers OpenAI-compatible access, model routing, enterprise uptime claims, and a set of priced models including GLM, DeepSeek, Kimi, and MiniMax. |
Developer Tools / AI API | 81 | ↑ +2 | 15 days ago | Details |