AgentDish directory
ROCm
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#95
→ 0
vLLM v0.28.0
Release page for vLLM v0.28.0, a high-throughput and memory-efficient inference and serving engine for LLMs. The snapshot shows release highlights, model support updates, breaking changes, and installable artifacts for PyPI, Docker, ROCm, CPU, and XPU. |
Developer Tools / LLM Serving / Inference | 89 | → 0 | 10 days ago | Details |
|
#1236
↓ -2
Speculative Decoding in vLLM on AMD GPUs
A detailed vLLM blog post explaining speculative decoding on AMD GPUs, with mechanics, supported drafting methods, configuration guidance, tuning advice, and benchmark discussion. |
AI/ML / Inference Optimization | 82 | ↓ -2 | 2 days ago | Details |
|
#1604
↑ +2
tiny-vLLM AMD GPU support via HIP
A pull request adding AMD GPU support to tiny-vLLM through ROCm/HIP while keeping the existing CUDA build path unchanged. The snapshot describes the compatibility header, CMake option, architecture selection, and validation on multiple AMD GPUs. |
Developer Tool / LLM inference engine | 79 | ↑ +2 | 72 days ago | Details |