GitHub ↗
Inference

vLLM by Inferact

https://vllm.ai/

1 recorded version · last updated 2026-08-30

2026-08-30CURRENT

The High-Throughput and Memory-Efficient inference and serving engine for LLMs

Easy, fast, and cost-efficient LLM serving for everyone.

vLLM by Inferact hero section on 2026-08-30

More in Inference & Model Serving