AgentDish directory

serving

Accepted listings with this tag.

Listing Category Score Trend Checked
#78 → 0
vLLM v0.28.0

Release page for vLLM v0.28.0, a high-throughput and memory-efficient inference and serving engine for LLMs. The snapshot shows release highlights, model support updates, breaking changes, and installable artifacts for PyPI, Docker, ROCm, CPU, and XPU.

Developer Tools / LLM Serving / Inference 89 → 0 just now Details