AgentDish directory

ROCm

Accepted listings with this tag.

Listing Category Score Trend Checked
#95 → 0
vLLM v0.28.0

Release page for vLLM v0.28.0, a high-throughput and memory-efficient inference and serving engine for LLMs. The snapshot shows release highlights, model support updates, breaking changes, and installable artifacts for PyPI, Docker, ROCm, CPU, and XPU.

Developer Tools / LLM Serving / Inference 89 → 0 10 days ago Details

A detailed vLLM blog post explaining speculative decoding on AMD GPUs, with mechanics, supported drafting methods, configuration guidance, tuning advice, and benchmark discussion.

AI/ML / Inference Optimization 82 ↓ -2 2 days ago Details

A pull request adding AMD GPU support to tiny-vLLM through ROCm/HIP while keeping the existing CUDA build path unchanged. The snapshot describes the compatibility header, CMake option, architecture selection, and validation on multiple AMD GPUs.

Developer Tool / LLM inference engine 79 ↑ +2 72 days ago Details