Altern Altern
2027 AIs indexed
Home
vLLM

vLLM

The open-source inference engine most self-hosted LLM deployments run on

Free plan
Copied!
Visit

Author

Dariush Abbasi

Pricing

Free

Updated

1 day ago

Screenshots

/

About

vLLM is a high-throughput, memory-efficient open-source library for serving large language models, built around its PagedAttention algorithm. It underpins a large share of production and research LLM-serving infrastructure.

Sign in to continue

It's easier when you're signed in — Altern helps you get more out of AI.

By continuing you agree to our Terms and Privacy Policy.