Inference provider · Inference gateway (self-host)

LiteLLM

Verified Sep 2026

Quick answer: Self-hosted OpenAI-compatible gateway with routing, budgets, and observability across 100+ models.

Best for

Self-hosted OpenAI-compatible gateway with routing, budgets, and observability across 100+ models.

Skip if

You want a paste-key hosted aggregator (OpenRouter/Requesty) with zero gateway ops.

Who it fits

  • Self-hosted OpenAI-compatible gateway with routing, budgets, and observability across 100+ models.
  • Builders who start in the browser without a local IDE setup
  • Terminal-first engineers who live in the shell and git
Experimental

Pros

Self-hostable OSS proxy/gateway routing 100+ models through one OpenAI-compatible API — load balancing, fallbacks, budgets, and observability for teams that want control vs a hosted aggregator.

Cons

You operate the gateway (Docker/K8s) — not a paste-key SaaS like OpenRouter; model keys still BYOK per provider; catalog has not hands-on verified production ops.

Facts

Locality:
Cloud-native
Surfaces:
Web, CLI
Maturity:
Experimental
BYOK:
Bring your own API key
Open source:
Source available
Free tier:
Free option available
Verified:
Sep 2026

Agent standards & memory

AGENTS.md: N/ASKILL.md: N/ARules / memory: LiteLLM proxy config; provider API keys in gateway env.100% Free (Open-source core); pay underlying model providers.

Compare one neighbor first

Vendor pages sell hard. Read one alternative on this site, then leave for the official link if the fit still holds.

Alternatives

Related tools

Someone asks you about vibe-coding software? Share this site with them — we are adding more useful data regularly. agents.dancingteeth.net

← Back to full tool directory