Open Source AI

vLLM

High-throughput LLM serving

Open-source inference engine known for PagedAttention and production serving performance.

inferenceservingopen source

Overview

Open-source inference engine known for PagedAttention and production serving performance.

vLLM is listed in Axioniq’s curated Open Source AI collection. Open-source models, frameworks, and communities that keep the AI ecosystem inspectable and portable.

At a glance

  • Listed under Open Source AI.
  • High-throughput LLM serving
  • Tagged for inference.
  • Tagged for serving.
  • Tagged for open source.

Who it’s for

  • teams that prefer open models and portable stacks

Category context

Open-source models, frameworks, and communities that keep the AI ecosystem inspectable and portable.

Compare peers in Open Source AI or jump straight to the official vLLM site.