pip install vllm

Category: AI/ML Tooling

pip install vllm

Install the vLLM inference engine

Installs vLLM with CUDA support for serving LLMs, pulling in large dependencies like PyTorch, so expect a sizeable download. After install, `vllm serve` launches an OpenAI-compatible API server. CUDA must match your driver; vLLM wheels target recent NVIDIA driver versions.
Looking for more? Search all 7,657 commands — works offline, in English or Spanish, and fixes typos.