Skip to main content

Start free. Pay for what you use.

Relay bills on the tokens it routes, Sidekick on seats, and Visibility on the scope you connect. Model provider costs stay on your own accounts in every case.

Get a demo

Model routing that starts free and scales with traffic.

Relay sits in front of your model providers and routes each request to the best-fit model for quality, cost, and latency. Pricing is tied to routed usage and custom routing surfaces.

Pro

Your first 10M routed tokens and 3 routers are free every month. The $0.15 only applies to what you route past that.
$0.15/1M tokens routed

$99/custom router/month for routing setups beyond the included routers.

  • OpenAI-compatible endpoint
  • Savings and usage dashboard
  • Privacy-preserving routing mode
  • Bring your own model keys
Start for free

Enterprise

Custom

Volume pricing, contracts, security review, and support.

  • Volume discounts
  • Organization-wide analytics
  • SSO and advanced admin controls
  • SLAs and priority support
  • Platform packaging with Sidekick and Visibility
Get a demo

Model provider costs billed separately to your own accounts.

Everything included

What each plan actually gives you.

Inclusion
Pro
Enterprise
Routing
Per-request model selection

Every call is matched to a model on quality, cost, and latency rather than pinned to one.

Included routed tokens

Tokens Relay routes each month before pay-as-you-go pricing starts.

10M / month
Volume terms
Custom routers

A router trained on your own traffic and scoped to one workload.

3 free, then $99 each
Volume terms
Fallback model

If a provider fails or times out, the request completes somewhere else.

Privacy-preserving routing

Routes on request metadata so prompt contents never need to leave your control.

Integration
OpenAI-compatible endpoint

Point an existing SDK at one base URL. No client rewrite.

Bring your own model keys

Provider calls run on your accounts, so provider spend is never resold.

Shadow mode

Run a candidate beside your primary on real traffic before anything switches.

Reporting and admin
Savings and usage dashboard

What was routed, what it cost, and what the alternative would have cost.

Organization-wide analytics

Usage rolled up across every team and workspace, not just your own.

SSO and advanced admin controls

Directory-backed sign-in with role and policy administration.

SLAs and priority support

Contracted response times and a named support path.

Relay questions.

Relay predicts which model should handle a request based on the workload, quality target, cost, latency, and routing policy you choose.

Relay charges for the routing layer: routed tokens and custom routers. The model calls themselves still run through your own provider accounts.

A custom router is a saved routing setup for a specific workload, model pool, objective, or environment. Pro includes the base routing path; additional custom routers are $99/router/month.

Relay is focused on the routing decision. It can sit behind or alongside gateway infrastructure, but the core job is deciding which model should handle each request.

Yes. Relay is designed for long-running agent workloads where model choice, context shape, latency, and cost change across the session.

Relay can route traffic for Claude Code-style workflows when your team wants model selection, cost control, or routing policy around those sessions.

Relay evaluates request shape, context needs, past outcomes, model pool, cost objective, latency target, and your routing policy before choosing the model path.

Savings depend on traffic mix and quality requirements. The usual goal is to move requests away from unnecessarily expensive models without degrading the output your workflow needs.

Yes. Relay can learn from routing outcomes and usage patterns so recommendations improve as more company traffic flows through it.

Privacy-preserving routing limits what the router needs to inspect by using derived request features and metadata where possible instead of exposing raw payloads broadly.

Enterprise Relay can include SSO, advanced admin controls, security review support, SLAs, and priority support as part of the commercial package.

Relay is designed to keep routing overhead small, but the exact impact depends on model pool, policy, privacy mode, and where Relay sits in your request path.

Relay has a free tier, then Pro pricing based on routed tokens plus custom routers. Enterprise pricing is available for volume, controls, SLAs, and platform packaging.

Build if routing is core infrastructure your team wants to maintain. Buy if you want the routing layer, continuous model tracking, analytics, and policy controls without dedicating a research and infra team to it.

Start with a narrow workload, connect the relevant model providers, define the model pool and objective, then compare quality, latency, and cost before expanding traffic.

No. Provider keys are managed by admins and used server-side. Employees and app users do not see the keys.