One unified endpoint
Use one OpenAI-compatible API instead of maintaining separate SDKs, credentials, and request formats for each provider.
Access AI models through one familiar API—without rebuilding your integration for every provider.
1# Official OpenAI SDK - only base_url changes2from openai import OpenAI34client = OpenAI(5 base_url="https://www.realrelay.ai/v1",6 api_key="sk-realrelay-***",7)89response = client.chat.completions.create(10 model="deepseek-v4-flash",11 messages=[{"role": "user", "content": "Summarise this report"}],12)1314print(response.choices[0].message.content)Recently released models available in our catalog.
openai/gpt-6.1-solanthropic/claude-sonnet-5-5anthropic/claude-opus-5-5openai/gpt-6-solopenai/gpt-6-lunaopenai/gpt-6-astraanthropic/claude-fable-5-1z-ai/glm-5.3deepseek/deepseek-v4.1-flashz-ai/glm-5.3-flashopenai/gpt-image-2.5-sunburstopenai/gpt-image-2.5-flareWhy RealRelay.ai
Replace separate provider integrations with one API key and one familiar request format, with clear model choices, published prices, and routing options.
Use one OpenAI-compatible API instead of maintaining separate SDKs, credentials, and request formats for each provider.
Compare published model prices in one catalog, then use one prepaid credit balance across available models.
Keep the familiar request format and move between supported models without rewriting your integration.
When a model has multiple channels, choose the default, lowest-price, or lowest-latency strategy at the account level.
Credit tiers
Top up when you need to, then use credits across the connected model catalog without a monthly commitment.
Support higher-volume API usage, batch processing, and frequent automation.
Top-up amount
$500
Fund large model workloads and business-critical production applications.
Top-up amount
$1,500
Credit-based, not a monthly subscription. Your balance is deducted only when you use it.
01
What the gateway is and how an existing integration moves over.
No. RealRelay.ai is not a model provider. Instead, it connects to upstream model providers to offer reliable, stable, secure, and unified smart routing and model access. It supports multiple API formats compatible with mainstream providers, including OpenAI, Anthropic, and Gemini.
Usually not. Point base_url at https://www.realrelay.ai/v1, swap in a RealRelay.ai key, and pick a model id from the catalog. The request shape stays the same.
From a catalog snapshot the operator publishes from the platform — the model page shows the date it was published. Struck-through original prices come from our channel list-price records; what we charge always comes from the platform.
02
How channel choice, model capabilities, and usage charges work.
Yes. Write the bare model name to let the gateway pick the cheapest channel serving it, or prefix it with a channel name to lock a specific one. The routing strategy (default, lowest price, lowest latency) is an account-level setting in the console.
Streaming works via the standard stream parameter over SSE. Tool calling depends on the model. The catalog lists which models report it.
You prepay credits in the console wallet, and usage is deducted per million tokens at the published rate, at direct provider cost without hidden markups. Image and media models bill per call. The catalog shows the mode alongside the price.