Experiential
9.6k

Pricing

Credits are the unit. Routed usage, intelligence features, and plan allotments all bill in them, and a credit is worth the same everywhere. Start free, pick Pro, Max or Ultra (each step up raises your limits and makes models cheaper), or scale on Enterprise. Routed tokens stay at provider cost, 0% on top.

Free

$0per month

500 credits a month once you verify a card (a one-time $1 charge, added to your balance). Route on credits at each provider’s list price with nothing on top.

  • Every hosted provider, one endpoint
  • 0% markup on routed tokens
  • Budgets, allowlists, attribution
  • Community support

Pro

Most popular
$20per month
  • Everything in Free
  • Actionable intelligence
  • Bring your own key & local models
  • Prompt and response storage off
  • Zero data retention
  • Tier 2 limitsFree Tier 1: 240 requests/min, 4M tokens/minPro Tier 2: 1,500 requests/min, 25M tokens/minMax Tier 3: 6,000 requests/min, 100M tokens/minUltra Tier 4: 10,000 requests/min, 150M tokens/minNeed more than Ultra? Tier 5 is by request.
  • Free DeepSeek V4.1 FlashApplies to your first 2,000 credits of usage each month, then list price.
  • Free GPT-6 LunaApplies to your first 2,000 credits of usage each month, then list price.
  • Free JevApplies to your first 2,000 credits of usage each month, then list price.
  • Up to 50% off other models50% off DeepSeek V4 Flash, Qwen3.8 27B11.25% off Kimi K35% off GPT-4.1, GPT-4.1 Mini, GPT-4.1 Nano, GPT-4o, GPT-4o Mini, GPT-5, GPT-5 Mini, GPT-5 Nano, GPT-5.1, GPT-5.2, GPT-5.4, GPT-5.4 Mini, GPT-5.4 Nano, GPT-5.5, GPT-5.6 Luna, GPT-5.6 Sol, GPT-5.6 Terra, GPT-6 Astra3.75% off Qwen3.8 Max2.5% off Claude Fable 5, Claude Fable 5.1, Claude Sonnet 5, GLM-5.3 Flash, Qwen3.8 2.4T A95B, Qwen3.8 Flash, Tencent Hy4 Preview1.25% off Claude Haiku 4.5, Claude Opus 4.5, Claude Opus 4.6, Claude Opus 4.7, Claude Opus 4.8, Claude Sonnet 4.5, Claude Sonnet 4.6Applies to your first 2,000 credits of usage each month, then list price.

Enterprise

Customscoped with your team

The managed layer around the gateway: identity, policy, data residency, and the intelligence engagement, scoped with your team at the lowest per-credit rate.

Book a call
  • Committed credits at the lowest rate
  • Slack support
  • SSO: SAML and SCIM provisioning
  • Advanced RBAC and org-wide policies
  • Private networking and data residency
  • Security reviews and SOC 1 access
  • Red carpet support
Every plan includes
FreePro · Max · UltraEnterprise
Plans and credits
Included credits per month500 (after $1 card verification)2,000 to 1,000,000Custom
Free models (Pro and up; Ultra adds more)Not includedDeepSeek V4.1 Flash, GPT-6 Luna, JevCustom
Discounts on other models (Pro / Max / Ultra)Up to 25%Up to 50% / 75% / 45%Custom
Rate limitsTier 1Tier 2 / 3 / 4Tier 5 and custom
Early releasesNot includedMax and UltraIncluded
Flat 1¢ per credit, 0% token markupIncludedIncludedIncluded
Buy more credits anytime (top-ups)IncludedIncludedIncluded
Markup on routed tokens0%0%0%
One bill across providersIncludedIncludedIncluded
Gateway
OpenAI-compatible /v1 endpointIncludedIncludedIncluded
Every hosted providerIncludedIncludedIncluded
Your own provider keys (BYOK)Not includedIncludedIncluded
Local and custom modelsNot includedIncludedIncluded
Provider failoverIncludedIncludedIncluded
Controls
Budgets and hard capsIncludedIncludedIncluded
API key managementIncludedIncludedIncluded
Model allowlists per key or agentIncludedIncludedIncluded
DashboardIncludedIncludedIncluded
Observability
Usage by agent, person, model, dayIncludedIncludedIncluded
Request logs with route and costIncludedIncludedIncluded
Live traffic viewIncludedIncludedIncluded
Security
Route only to ZDR providersIncludedIncludedIncluded
No-training provider policyIncludedIncludedIncluded
Provider allowlistsIncludedIncludedIncluded
Provider keys never reach agentsIncludedIncludedIncluded
Intelligence
Per-prompt model optimizationNot includedIncludedIncluded
CachingNot includedIncludedIncluded
Intelligence features, billed in creditsNot includedIncludedIncluded
A model you own, trained on your trafficNot includedNot includedIncluded
Support
Community (Discord and GitHub)IncludedIncludedIncluded
Dedicated supportNot includedIncludedIncluded
Slack supportNot includedNot includedIncluded
Security reviewsNot includedNot includedIncluded

Routing policies, budgets, and allowlists are enforced at the gateway on every request. No code changes.

Enterprise-ready.

The posture an enterprise rollout asks for: what ships today, what we scope with your team, and what is on the roadmap. For the intelligence layer, security reviews, and dedicated support, talk to us.

Book a call
Available today

Multiple organizations

Separate orgs, keys, and billing under one account.

Google and GitHub sign-in

OAuth sign-in for the whole team, no passwords required.

Usage history and attribution

Usage, spend, requests, TTFT, and token counts at key and team scope.

Usage API

Spend by agent, person, model, or day, pulled into your own systems.

API key management

Create, view, and delete keys from the dashboard or the API.

Zero Data Retention routing

Route only to providers under a ZDR agreement, enforced in policy.

No training on your data

Route only to providers that will not train on customer data.

Provider resilience

Failover across providers and pooled accounts on capacity errors.

Request logs

Every request recorded with its route, tokens, and cost.

Budgets and policies

Hard caps and model allowlists, enforced on every request.

SOC 1

SOC 1 compliance, report available on request.

Every release in the open

Read the code your traffic flows through, on GitHub.

View the repo ↗
On request

SSO: SAML and SCIM

Enterprise identity providers and directory sync.

Short-lived credentials

Expiring tokens in place of static keys.

Advanced RBAC

Fine-grained roles beyond the built-in org roles.

Org-wide policy packs

One policy set applied across every team and key.

Approval workflows

Changes reviewed before they land, with change history.

Managed KMS

Customer-managed keys for stored provider credentials.

Private networking

Private connectivity between your VPC and the gateway.

Data residency

Regional routing and storage constraints.

Enterprise support

A named contact and a shared channel.

Planned

Managed upgrades and backups

For self-hosted deployments we operate with you.

Managed optimization workers

Dedicated capacity for the intelligence layer.

Forecasting and alerts

Spend projections and threshold notifications.

Questions

How do credits work?

A credit is one standardized unit of usage. Routed tokens spend credits at the provider's list price, and the intelligence features spend them too: one unit across everything, worth the same wherever it is spent. Plans and top-ups are just how you buy credits.

Is there really no markup on tokens?

Yes. A routed token spends credits at the provider's list price. We add $0.00 per token, on every plan. What you pay for is the credits themselves: the plan you pick and the intelligence layer, never a percentage of your traffic.

Plans vs. top-ups?

A plan is a monthly credit allotment at a flat 1¢ per credit, the same rate as a top-up. Top-ups let anyone buy credits without a plan. What a plan adds is the unlocks: BYOK and local models, actionable intelligence, prompt storage off, zero data retention, higher rate limits, and discounts on models. $20 to $199 a month is Pro, $200 to $1,999 is Max, $2,000 and up is Ultra; each step up raises your limits and deepens the discounts, and the discounts cover your plan's monthly credits of usage.

What happens when my credits run out?

Requests on credits stop at your caps, so nothing silently overruns. Add more from the dashboard, and anything routed through your own provider keys keeps working the whole time.

Can I bring my own provider keys?

Yes, on Pro, Max, Ultra and Enterprise. The gateway routes through your keys first, so existing provider commitments and discounts flow through untouched, and traffic on your own keys is never rate limited by us.

What does self-hosting include?

The whole gateway: the OpenAI-compatible /v1 endpoint, key management, budgets, provider waterfalls, and the usage API. uvx --from experiential exp run starts it on your own infrastructure.

Get started

Route your first request today.

One base URL swap. Provider prices, hard budgets, every token attributed.

Book a call