Kembali ke katalog
LLM RunnerPythonApache-2.0
vLLM
vllm-project/vllm
High-throughput and memory-efficient inference and serving engine for LLMs.
Install Command
pip install vllm
GitHub Stats
Stars91,073
Forks21,784
LanguagePython
Deployment
✔ Linux (native / Docker)
✔ macOS
✔ Windows (WSL2)
✔ Cloud (AWS · GCP · Azure)