99.9%
Gateway requests · monthly measurement window
Monthly availability target for gateway requests.
Access GPT, Claude, Gemini and 50+ models through one OpenAI-compatible API, with configurable provider fallback, usage controls and request-level visibility.
One API for leading model providers
Three-line integration
RouteAPI uses an OpenAI-compatible interface, so most integrations can keep the existing SDK and update the API key and base URL.
Supported API features may vary by model and upstream provider.
Read the API docsMULTI-MODEL ACCESS
Send requests through one OpenAI-compatible endpoint while selecting the supported model your application needs.
MODEL ROUTING & RESILIENCE
Choose a model and configure its primary and fallback providers. RouteAPI monitors channel health and routes requests through the available path according to your policy.
Simulate a provider failure to see the next request move to the configured fallback.
Your application selects a model and sends requests through RouteAPI. You configure the provider order; RouteAPI monitors channel health and uses the configured primary or fallback path according to your policy. A provider failure affects the next request, not an in-progress streaming response.
Transparent pricing
25% below official API list prices
On our most-used US-based models
20% below official API list prices
On our most-used China-based models
No additional platform fee on top of the displayed model rate.
Billing principles: Pay only for billable usage from a completed request. A failed request with no billable usage is not charged; when a configured fallback completes the request, that usage is billed at the selected model's displayed RouteAPI rate.
Control plane
Replace separate provider endpoints, credentials and operational records with one OpenAI-compatible control layer.
Incoming requests
0
+12.4%Today's tokens
0
TodaySuccess rate
0.00%
Cache hit rate 86.4%Estimated cost
$0.00
Saved $36.79RELIABILITY BASELINE
Public reference targets for the gateway layer. Model inference and streaming generation are measured separately because their latency varies by workload.
99.9%
Gateway requests · monthly measurement window
Monthly availability target for gateway requests.
≤ 400 ms
RouteAPI processing only · model inference excluded
99% of eligible requests should add no more than 400 ms at the gateway layer.
≤ 1 s
Eligible provider or transport failures · next request
Internal recovery target; the objective is policy-defined because no universal failover time applies to every workload.
30+
Supported upstream families · current product scope
Product coverage count, shown separately from availability and latency SLOs.
These figures are our service level targets and current product scope, not live production telemetry. Contractual terms are agreed per enterprise plan.
DATA HANDLING TRANSPARENCY
A concise view of where requests go and what RouteAPI records. Upstream providers apply their own data policies.
Your application
Sends prompts and request parameters
RouteAPI gateway
Routes the request and records operational data
Configured upstream provider
Processes the request under its own data policy
Requests pass through RouteAPI and are sent to the upstream provider configured by the system for the selected model.
RouteAPI does not store your request content and will not use it for training. Your data is forwarded directly to the selected upstream provider; please refer to each provider's data policy.
Operational records include model, provider, status, latency, token usage and cost.
This page describes RouteAPI's current request path and logging behavior; it does not make a blanket guarantee for upstream retention, deletion or training policies.
Questions, answered
Five practical answers for your first production call. Deeper details stay in the docs and pricing catalog.
Usually not. Keep the OpenAI client, update the Base URL and API key, and check the compatibility matrix because supported features can vary by model and upstream provider.
Yes. You select the model you want to use, and RouteAPI automatically chooses the most suitable provider based on the system-configured route and provider status.
You configure a primary provider, fallback providers and retry policy for a model route. RouteAPI applies those rules when a qualifying provider or transport failure occurs; not every error can be recovered.
Billing follows billable usage. A failed attempt with no billable usage is not charged; a retry or fallback that generates billable usage can be charged, and a fallback that completes the request is billed at the selected model rate. Read billing details
No. RouteAPI does not store prompt or response content. For billing, we retain only necessary metadata such as the model, token usage, cost, request status and timestamp.
Create an API key, update your base URL and send a test request.
Pay as you go · Save up to 25%
Trusted at scale
The teams building on RouteAPI
A production-ready gateway for teams that need reliable access to the models behind their products.