vllm-project / vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

维护停滞 Apache 2.0 Python Notable
90.4k 21.4k 41 分钟前
CIPyPI
gpt llm pytorch model-serving transformer llm-serving inference llama amd cuda tpu deepseek qwen blackwell deepseek-v3 gpt-oss kimi moe openai qwen3

星标趋势

数据积累中,暂无足够数据生成趋势图

AI 分析

项目摘要

vllm-project/vllm: A high-throughput and memory-efficient inference and serving engine for LLMs

为什么值得关注

High popularity with 90,438 stars, active contributor community (3014 contributors), frequent releases.

优势

  • Large and active community
  • Continuous integration configured
  • Test suite present

局限性

  • Limited documentation
分析模型:fallback | 分析时间:刚刚