AgentDish directory
openai-compatible-api
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#7
↑ +298
WebLLM
WebLLM is a high-performance in-browser LLM inference engine that runs locally in the browser with WebGPU acceleration. It exposes an OpenAI-compatible API, supports streaming and JSON mode, and includes examples for building chat apps and browser extensions. |
Developer Tool / AI SDK / In-browser LLM inference | 92 | ↑ +298 | 117 days ago | Details |
|
#134
→ 0
fusionHarness
An open-source mixture-of-agents server that fans prompts out to multiple LLMs, judges the responses, and synthesizes a final answer. It exposes an OpenAI-compatible API and also includes an agentic coding harness plus integrations for tools like Claude Code, aider, Continue, LangChain, and Pi. |
Developer Tools / AI Orchestration | 89 | → 0 | 75 days ago | Details |
|
#160
↓ -3
GIDE
GIDE is a privacy-first AI coding command center that runs local or cloud models in a code editor and terminal. The page shows live session examples, a write-approval gate, CLI support, and an OpenAI-compatible local API. |
Developer Tools / AI Coding Assistant | 88 | ↓ -3 | just now | Details |
|
#330
↓ -4
SIE: Superlinked Inference Engine
Open-source self-hosted inference server for agent workloads, with an OpenAI-compatible API, multi-model serving, and production cluster tooling. |
Developer Tools / AI Infrastructure | 87 | ↓ -4 | 21 days ago | Details |
|
#380
↓ -4
Wyolet Relay
Wyolet Relay is a self-hosted LLM router that exposes OpenAI- and Anthropic-compatible endpoints in front of multiple providers. It pools API keys, handles failover and rate limits, tracks usage and cost, and includes a quickstart plus Docker-based deployment. |
Developer Tools / API Infrastructure / LLM Routing | 87 | ↓ -4 | 73 days ago | Details |
|
#559
↑ +2
ZSE v2.0.0
A pure-Python LLM inference engine and server with CUDA/HIP/Metal code generation, OpenAI-compatible API support, built-in RAG, and multi-GPU backend support. |
Developer Tools / AI / ML Infrastructure | 86 | ↑ +2 | 90 days ago | Details |
|
#850
↓ -6
SayItDev
A macOS command-line tool for on-device AI and voice tasks, including speak, listen, transcribe, chat, and an OpenAI-compatible local server. |
Developer Tool / CLI / Local AI | 84 | ↓ -6 | 49 days ago | Details |
|
#1433
↓ -1
How to Setup a Local Coding Agent on macOS
A detailed walkthrough for running a local coding agent on macOS with llama.cpp, Gemma 4, MTP speculative decoding, image support, and Pi as the agent interface. |
Developer Tools / Code Assistant | 80 | ↓ -1 | 79 days ago | Details |
|
#1531
↓ -208
Bonsai 1.7B: Apple Silicon Optimized Build
An Apple Silicon–optimized inference build of Bonsai 1.7B with custom Metal kernels, benchmark results, quick-start instructions, and a bundled OpenAI-compatible server. |
Developer Tools / Code Assistant | 79 | ↓ -208 | 117 days ago | Details |
|
A vLLM blog post explaining Semantic Router as an open serving-layer runtime for bounded micro-agents, with routing patterns like Confidence, Ratings, ReMoM, Fusion, and Workflows behind a single model API. |
Developer Tools / Code Assistant | 76 | ↓ -1 | 62 days ago | Details |