LiteLLM Proxy gateway details

An open-source LLM gateway/proxy layer that unifies many providers into an OpenAI-style interface; ideal for teams that want to self-host a relay and control their own keys and logs.

Small amounts
Trial first

Overview

An open-source LLM gateway/proxy layer that unifies many providers into an OpenAI-style interface; ideal for teams that want to self-host a relay and control their own keys and logs.

Best for: For dev teams with ops capacity who want control over API keys, logs, routing, and permissions.

Availability and latency

Current status
Self-hosted
Latency reference
Depends on your deployment
/models check
Self-hosted configuration
Last verified
2026-07-01
Status notes
An open-source proxy layer; real latency and availability depend on your deployment and upstream providers.

Supported models

OpenAIClaudeGeminiDeepSeekGrokQwenMoonshotDoubaoEmbedding

Pricing and multipliers

Pricing mode
Open-source and self-hosted; cost comes from servers and the underlying model providers.
Price tags
High stability, Self-hosted
Free quota
No unified free quota; depends on the providers you connect.

Payment methods

Self-hosted

How to call the API

Usually you create an API key in the platform console, then replace the SDK's baseURL, model name, and auth details per the docs. Test streaming, function calling, error codes, and rate limits before launch.

import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.AI_GATEWAY_API_KEY,
  baseURL: "see the platform docs"
});

const response = await client.chat.completions.create({
  model: "gpt-4.1-mini",
  messages: [{ role: "user", content: "Test one call through an AI gateway." }]
});

Pros

  • Supported models: OpenAI, Claude, Gemini, DeepSeek, Grok, Qwen, Moonshot, Doubao and more.
  • Capabilities: OpenAI-compatible, Streaming, Function calling, Embeddings, Log query.
  • Documentation completeness: High; editorial stability score 4.0 / 5.

Cons

  • Self-hosting means taking on ops, security, rate-limiting, and key management; the learning curve is high for individuals.
  • Third-party relay pricing, models, and availability can change, so re-check them periodically.

Alternatives

OpenRouter

Details
Aggregator

A multi-model routing platform for developers; switch models by changing the base_url with the OpenAI SDK, ideal for teams comparing the quality and price of many models at once.

Models
OpenAI, Claude, Gemini, DeepSeek, Grok, Qwen and more
Payment
Credit card, Stripe, Free trial
Price
Billed per model and provider; some models offer free routing or free quota.

Vercel AI Gateway

Details
Cloud vendorHigh stability

Vercel's official AI Gateway, emphasizing unified model calls, failover, caching, and observability; ideal for apps already deployed on Vercel.

Models
OpenAI, Claude, Gemini, DeepSeek, Grok, Qwen and more
Payment
Credit card, Free trial
Price
Settled via your Vercel account and model usage; the docs stress the gateway itself adds no markup.

302.AI

Details
Aggregator

A tool and API aggregation platform covering text, image, video, and audio AI capabilities; ideal for individuals and small teams wanting one-stop access to many models.

Models
OpenAI, Claude, Gemini, DeepSeek, Grok, Qwen and more
Payment
Credit card, Free trial
Price
Billed per tool or model usage; the site lists separate prices per model.

OhMyGPT

Details
AggregatorPrepaidLow cost

A common API relay service among Chinese users; the site highlights aggregation of OpenAI, Claude, and Gemini plus multiple top-up methods.

Models
OpenAI, Claude, Gemini, DeepSeek, Grok, Qwen and more
Payment
Alipay, WeChat, USDT, PayPal, Stripe, Free trial
Price
Pay-as-you-go after top-up; the site lists per-model token prices.

FAQ

Is LiteLLM Proxy compatible with the OpenAI API?
This entry is marked as supporting the OpenAI-compatible format, but the exact base_url, model names, and parameters still follow the official docs.
Is LiteLLM Proxy ready for production?
Run a small-traffic load test, rate-limit test, and error-code test first, and confirm the billing rules; keep a backup provider for important workloads.

Risk

Risk notice

This site only organizes information about AI API relay services; it makes no guarantee about third-party platforms' stability, price, data security, or availability. Judge the risk yourself before use, and avoid uploading sensitive data, personal privacy, trade secrets, or core API keys to unknown platforms.