BYOAIK (Bring Your Own AI Key) is a directory of AI tools that run on your own API key.

vLLM

High-throughput self-hosted serving engine for open models on your own GPUs, behind an OpenAI-compatible API you own.

Website Source code

Category:
Self-Hosted AI Platforms
Maintenance:
Actively developed (last commit today)
Pricing:
Open Source
Open source:
Yes (Apache-2.0)
Self-hostable:
Yes
Local-first:
Yes
Platforms:
Self-hosted, Docker, Library/SDK
AI providers (bring your own key):
Ollama, Hugging Face, Custom / OpenAI-compatible
API key storage:
No cloud key required for local model
Key risk level:
LOW
Trust score:
87/100

vLLM is an inference and serving library for running open-weight models on your own NVIDIA, AMD or Intel GPUs at production throughput, using PagedAttention and continuous batching. It exposes an OpenAI-compatible API server deployed on hardware you provision, so no vendor key is involved. It is widely used as the serving layer beneath self-hosted deployments and by inference providers that let you bring your own weights. Apache-2.0, 89,119 stars, pushed the day of review with very active development.

Why this trust score (87/100)

Trust measures how the tool treats your API key and how much of that has been verified. It contains no popularity signal.

  • Key Safety 25/25: No cloud key is needed at all. The model runs locally.
  • Request Routing 17/20: Requests go straight from you to the AI provider.
  • Transparency 20/20: Source is public under Apache-2.0, so anyone can check how the key is handled.
  • Privacy 14/15: Local-first: it works without sending your data anywhere. Can be self-hosted, so the data path stays inside infrastructure you control.
  • Maintenance 10/10: Actively developed: commits within the last three months.
  • Verification Confidence 1/10: Compiled from public documentation by an AI-assisted pass, not independently confirmed.

What was checked

Verification tier RESEARCH_ASSISTED, derived from the evidence below and not set by hand.

  • [STRONG · SOURCE_SCAN] Supports a local model backend, so it can run with no cloud provider key at all. source
  • [CONFIRMED · SOURCE_SCAN] Most recent commit 2026-08-15, about 0 month(s) ago. source

How your API key is handled

Key handling varies. Requests are sent directly to the AI provider. Because it can be self-hosted, your key never has to touch a third-party backend.

Setup

Add your provider API key in settings and choose a model.