What is an LLM gateway (and which to use)?

An LLM gateway is a single, unified API that sits in front of many model providers so you use one key and one base URL for all of them, with routing, failover and billing handled for you. UberLLM is an OpenAI-compatible LLM gateway across 380+ models that routes each request to the cheapest healthy provider, adds an open seller marketplace, and gives new accounts free credits.

base_url = "https://api.uberllm.dev/v1"

Live pricing

ModelInput / 1MOutput / 1M
Tencent HY3 (Free)$0.0000$0.0000
Seedream 4.5$0.0000$0.0000
FLUX.2 Flex$0.0000$0.0000
FLUX.2 Klein 4B$0.0000$0.0000
Z-Image Turbo (Venice)$0.0000$0.0000
Venice SD35$0.0000$0.0000
Flux 2 Pro (Venice)$0.0000$0.0000
Flux 2 Max (Venice)$0.0000$0.0000
Grok Imagine (Venice)$0.0000$0.0000
Grok Imagine High Quality (Venice)$0.0000$0.0000
Grok Imagine Pro (Venice)$0.0000$0.0000
GPT Image 1.5 (Venice)$0.0000$0.0000

Live prices via UberLLM · full pricing table →

FAQ

Why use an LLM gateway?

It removes per-provider contracts and keys, gives automatic failover when a provider is down, and lets you switch or compare models with one line of code.

Which LLM gateway is cheapest?

UberLLM routes to the cheapest healthy provider per request with no inference markup, so you always pay the lowest available price across providers.