Kembali ke katalog
LLM RunnerPythonApache-2.0

vLLM

vllm-project/vllm

High-throughput and memory-efficient inference and serving engine for LLMs.

Install Command
pip install vllm
GitHub Stats
Stars91,073
Forks21,784
LanguagePython
Deployment
✔ Linux (native / Docker)
✔ macOS
✔ Windows (WSL2)
✔ Cloud (AWS · GCP · Azure)