One endpoint forevery model.

Access AI models through one familiar API—without rebuilding your integration for every provider.

1# Official OpenAI SDK - only base_url changes2from openai import OpenAI34client = OpenAI(5    base_url="https://www.realrelay.ai/v1",6    api_key="sk-realrelay-***",7)89response = client.chat.completions.create(10    model="deepseek-v4-flash",11    messages=[{"role": "user", "content": "Summarise this report"}],12)1314print(response.choices[0].message.content)
164 models · 10 vendors · Catalog updated Oct 4, 2026

Why RealRelay.ai

Stop rebuilding AI infrastructure from scratch

Replace separate provider integrations with one API key and one familiar request format, with clear model choices, published prices, and routing options.

01

One unified endpoint

Use one OpenAI-compatible API instead of maintaining separate SDKs, credentials, and request formats for each provider.

02

Published prices, one credit balance

Compare published model prices in one catalog, then use one prepaid credit balance across available models.

03

Change the model ID, not your application

Keep the familiar request format and move between supported models without rewriting your integration.

04

Routing you can control

When a model has multiple channels, choose the default, lowest-price, or lowest-latency strategy at the account level.

Credit tiers

Choose the right credit tier

Top up when you need to, then use credits across the connected model catalog without a monthly commitment.

Study

01

Explore models, compare outputs, and test an initial API integration.

Top-up amount

$20

Standard

02

Run daily development, evaluation, and application workloads.

Top-up amount

$100

Most popular

Pro

03

Support higher-volume API usage, batch processing, and frequent automation.

Top-up amount

$500

Max

04

Fund large model workloads and business-critical production applications.

Top-up amount

$1,500

Credit-based, not a monthly subscription. Your balance is deducted only when you use it.

Questions

FAQ

The practical details to know before your first production request.

View all FAQs

01

Getting started

What the gateway is and how an existing integration moves over.

Is RealRelay.ai a model provider?

No. RealRelay.ai is not a model provider. Instead, it connects to upstream model providers to offer reliable, stable, secure, and unified smart routing and model access. It supports multiple API formats compatible with mainstream providers, including OpenAI, Anthropic, and Gemini.

Do I have to rewrite an existing OpenAI integration?

Usually not. Point base_url at https://www.realrelay.ai/v1, swap in a RealRelay.ai key, and pick a model id from the catalog. The request shape stays the same.

Where does the catalog data come from?

From a catalog snapshot the operator publishes from the platform — the model page shows the date it was published. Struck-through original prices come from our channel list-price records; what we charge always comes from the platform.

02

Routing and billing

How channel choice, model capabilities, and usage charges work.

Can I pin a specific channel?

Yes. Write the bare model name to let the gateway pick the cheapest channel serving it, or prefix it with a channel name to lock a specific one. The routing strategy (default, lowest price, lowest latency) is an account-level setting in the console.

Are streaming and tool calling supported?

Streaming works via the standard stream parameter over SSE. Tool calling depends on the model. The catalog lists which models report it.

How is billing structured?

You prepay credits in the console wallet, and usage is deducted per million tokens at the published rate, at direct provider cost without hidden markups. Image and media models bill per call. The catalog shows the mode alongside the price.