vLLM
High-throughput inference server for self-hosting open models.
- Category
- Code & Dev
- Pricing
- Free
- Website
- docs.vllm.ai
Tags
vLLM compared
Alternatives to vLLM
Aider
Free
Pair programming in your terminal that commits as it goes.
Cline
Free plan
Autonomous coding agent living inside VS Code.
Continue
Free
Open-source IDE assistant you point at any model, including local.
fal
Paid
Fast hosted inference for image, video and audio models.
Groq
Free plan
Inference on custom silicon — startlingly fast token throughput.
Hugging Face
Free plan
The public registry for open models, datasets and demo Spaces.
Appears in
Work at vLLM? Grab your listing badge — free, and it links back here.
Listing details are checked by hand but vendors change pricing without warning — confirm on vLLM's own site before buying. Some outbound links earn us a commission at no cost to you; see the disclosure.