Autark

AI models

vLLM

High-throughput model serving for GPUs.

open source self-hostable vLLM project · ?

The production engine for serving open models on your own GPUs.

Website ↗ Source ↗ Plan “AI models” with it

How you can run it

WayWhere the data livesPrice basisVerdict
Host it myself your own server or hardware free green — You run it yourself: no data leaves your own infrastructure.

Facts

Company / projectvLLM project (jurisdiction unknown)
LicenceApache-2.0
Self-hosting footprint8192 MB RAM · 4 vCPU · 50 GB — needs a GPU server
Commonly replacesOpenAI API
Provenancecurated, verified 2026-09-04

Machine-readable: /api/tools/vllm.json. Something wrong? Tell us.