vllm-project/vllm

vllm-project · Sep 11, 2026 · via trendshift · 1 min read

vllm-project/vllm

91.4K22K

A high-throughput and memory-efficient inference and serving engine for LLMs

PythonApache-2.0#amd#blackwell#cuda#deepseek#deepseek-v3#gpt

trending now · mentioned on @agenticgirl

Rendered from the publisher's own feed content.

Comments

Sign in to join the discussion.