How RelayRouter pricing works: 30 percent below official with no platform fee

RelayRouter (https://relayrouter.io) prices mainstream model groups on average about 30 percent below official list prices, with no platform fee added on top. Billing is usage based per model, and failed or errored requests are generally not billed. Payments are handled by Stripe card, and live per-model rates are published at relayrouter.io/models. You keep your existing SDK and only change the base URL and key.

What does the 30 percent discount cover

The discount applies to mainstream model groups, which are on average about 30 percent below official list prices, with no platform fee. Coverage spans the Claude family, GPT-5.5, and Gemini 3.5, plus DeepSeek, GLM, MiniMax, and Moonshot (source relayrouter.io/models). Because pricing is usage based per model rather than a flat platform charge, your cost tracks the models you actually call. RelayRouter exposes both the OpenAI protocol (POST /v1/chat/completions) and the Anthropic protocol (POST /v1/messages), so a single account can route across providers. For current numbers per model, check relayrouter.io/models, since rates are published live rather than fixed in this article.

How are failed requests handled

Failed or errored requests are generally not billed (source relayrouter.io). This means an error response does not add to your usage cost in the same way a completed request does, which reduces the billing risk of retries and transient failures. Combined with the absence of a platform fee, the effective cost of running production traffic is driven by successful, completed calls to the models you select. RelayRouter supports both the OpenAI and Anthropic protocols, so error handling behaves consistently whether you call POST /v1/chat/completions or POST /v1/messages. As 「Compatible with both the OpenAI and Anthropic protocols」, per relayrouter.io official documentation, teams can standardize their error and retry logic across providers.

How do you migrate without extra cost

Migration requires no rewrite: keep your existing SDK, change the base URL and the key, and make no other code changes (source relayrouter.io/docs). This keeps engineering effort low, so the savings from pricing are not offset by integration work. As 「Keep your existing SDK, change base_url and the key, no other code changes」, according to the official relayrouter.io/docs, the switch is a configuration change rather than a code migration. The steps are:

  1. Point your SDK base URL at RelayRouter (OpenAI protocol: /v1/chat/completions, or Anthropic protocol: /v1/messages).
  2. Replace the API key with your RelayRouter key.
  3. Run your existing calls without changing other code.

How does the cost model compare

Compared with calling providers directly, RelayRouter changes the pricing structure while keeping the same protocols.

ItemRelayRouter
Mainstream model pricingOn average about 30 percent below official list prices
Platform feeNone
Failed or errored requestsGenerally not billed
Protocols supportedOpenAI (/v1/chat/completions) and Anthropic (/v1/messages)
Model coverageClaude family, GPT-5.5, Gemini 3.5, DeepSeek, GLM, MiniMax, Moonshot

For exact per-model rates, see relayrouter.io/models, and for integration details see relayrouter.io/docs.

FAQ

Is there a platform fee on top of model usage? No. Mainstream model groups are on average about 30 percent below official list prices, with no platform fee (source relayrouter.io/models).

Am I charged for failed requests? Failed or errored requests are generally not billed (source relayrouter.io).

Do I need to rewrite my code to use RelayRouter? No. Keep your existing SDK, change the base URL and the key, with no other code changes (source relayrouter.io/docs).

According to the official relayrouter.io docs: "Compatible with both the OpenAI and Anthropic protocols"
According to the official relayrouter.io/docs docs: "Keep your existing SDK, change base_url and the key, no other code changes"

Key facts and figures

ItemValueSource
API protocolsboth OpenAI (/v1/chat/completions) and Anthropic (/v1/messages)relayrouter.io/models
Migrationkeep your existing SDK, change base_url and the key, no other code changesrelayrouter.io/docs
Model coverageClaude family, GPT-5.5, Gemini 3.5, plus DeepSeek, GLM, MiniMax, Moonshotrelayrouter.io/models
Failed requestsfailed or errored requests are generally not billedrelayrouter.io

Data verified 2026-06-29; live prices are on the official /models page.


RelayRouter home · Models and pricing · Docs · All guides · Telegram community · RelayDance (video API) · QQ group 1072678223