vllm — tool profile and history

A high-throughput and memory-efficient inference and serving engine for LLMs

Active

Canonical source ↗

Language
Python
License
Apache-2.0
Created
Last activity
Topics
  • amd
  • blackwell
  • cuda
  • deepseek
  • deepseek-v3
  • gpt
  • gpt-oss
  • inference
  • kimi
  • llama
  • llm
  • llm-serving
  • model-serving
  • moe
  • openai
  • pytorch
  • qwen
  • qwen3
  • tpu
  • transformer

Current measurement

90 637GitHub stars+588 (+0.65 %)
measured growth

Momentum uses available readings from the last 7 days: measured with at least two comparable readings, estimated otherwise.

Reading history

90 days · 90 maximum readings
DateValueMetric
90 637GitHub stars
90 535GitHub stars
90 435GitHub stars
90 352GitHub stars
90 260GitHub stars
90 166GitHub stars
90 049GitHub stars
89 919GitHub stars
89 823GitHub stars
89 729GitHub stars
89 671GitHub stars
89 585GitHub stars
89 482GitHub stars
89 389GitHub stars
89 300GitHub stars
89 261GitHub stars
89 146GitHub stars
89 066GitHub stars
89 060GitHub stars
88 950GitHub stars
88 857GitHub stars
88 782GitHub stars
88 700GitHub stars
88 604GitHub stars

Classifications

Domains
Type
Other
Use cases
  • No declared use case

These entries declare the same topics. No similarity is inferred: only the shared topics are stated.