REPO //Trendshift
vllm-project/vllm_
A high-throughput and memory-efficient inference and serving engine for LLMs
★ 91.4K⑂ 22KPython
Sep 11, 2026[40]
https://github.com/vllm-project/vllm1 item tagged “amd” across every source.
A high-throughput and memory-efficient inference and serving engine for LLMs