Skip to content
V

vLLM

Inference and serving engine for large language models, built

Launch
#29in AI
–In stores

Description

Inference and serving engine for large language models, built for speed and hardware efficiency with an OpenAI-compatible API and support fo…

Preview

vLLM screenshot 1

Information

People also visit

Apps people often open alongside vLLM.

See all
O

Ollama

Load and run large LLMs locally to use in your terminal or build your apps

O

OpenHands

AI agent platform that runs autonomous coding agents to plan, write

C

Cherry Studio

Desktop AI client for Windows, macOS, and Linux

L

LiteLLM

Acts as a unified proxy across 100+ LLMs

L

LocalAI

Run LLMs, speech, image generation

P

Phoenix

Arize Phoenix: Open Source AI Development Platform

vLLM FAQ

What is vLLM?

Inference and serving engine for large language models, built for speed and hardware efficiency with an OpenAI-compatible API and support for a wide range of open models.

Do I need to download vLLM?

No. vLLM runs in your browser at vllm.ai, so there's nothing to install.

What are the best vLLM alternatives?

Popular alternatives to vLLM include Ollama, OpenHands and Cherry Studio. See all vLLM alternatives