Why it was accepted
The page clearly describes an AI infrastructure tool with a concrete use case: improving vLLM serving capacity through low-rank KV-cache compression. It includes install steps, a quick start, and measured results showing concurrency and context-length gains on real GPU hardware, which is enough for a useful public listing.