BYOAIK — Bring Your Own AI Key: a directory of AI tools that run on your own API key.

vLLM

High-throughput self-hosted serving engine for open models on your own GPUs, behind an OpenAI-compatible API you own.

Website Source code

Category:
Self-Hosted AI Platforms
Pricing:
Open Source
Open source:
Yes (Apache-2.0)
Self-hostable:
Yes
Local-first:
Yes
Platforms:
Self-hosted, Docker, Library/SDK
AI providers (bring your own key):
Ollama, Hugging Face, Custom / OpenAI-compatible
API key storage:
No cloud key required for local model
Key risk level:
LOW
Trust score:
87/100

vLLM is an inference and serving library for running open-weight models on your own NVIDIA, AMD or Intel GPUs at production throughput, using PagedAttention and continuous batching. It exposes an OpenAI-compatible API server deployed on hardware you provision, so no vendor key is involved. It is widely used as the serving layer beneath self-hosted deployments and by inference providers that let you bring your own weights. Apache-2.0, 89,119 stars, pushed the day of review with very active development.

Why this trust score (87/100)

Trust measures how the tool treats your API key and how much of that has been verified. It contains no popularity signal.

  • Key Safety 25/25 — No cloud key is needed at all — the model runs locally.
  • Request Routing 17/20 — Requests go straight from you to the AI provider.
  • Transparency 20/20 — Source is public under Apache-2.0, so anyone can check how the key is handled.
  • Privacy 14/15 — Local-first: it works without sending your data anywhere. Can be self-hosted, so the data path stays inside infrastructure you control.
  • Maintenance 10/10 — Actively developed — commits within the last three months.
  • Verification Confidence 1/10 — Compiled from public documentation by an AI-assisted pass, not independently confirmed.

What was checked

Verification tier: RESEARCH_ASSISTED — derived from the evidence below, not set by hand.

  • [STRONG · SOURCE_SCAN] Supports a local model backend, so it can run with no cloud provider key at all. source
  • [CONFIRMED · SOURCE_SCAN] Most recent commit 2026-08-15 — about 0 month(s) ago. source

How your API key is handled

Key handling varies. Requests are sent directly to the AI provider. Because it can be self-hosted, your key never has to touch a third-party backend.

Setup

Add your provider API key in settings and choose a model.