Skip to main content

The model

Everything that moves a request is free on every plan — routing across every provider, live model swapping, failover chains, spend caps, and bringing your own keys at zero markup. The subscription pays for the intelligence layer: the machinery that judges, compares, and proves your AI works on your real traffic. We never mark up tokens, never charge per seat, and never bill overages.

Plans at a glance

Rate limits exist on every plan as abuse prevention — they scale with your tier and are not something we sell. Units never gate routing: on Free they meter tracked traffic, on paid plans they size intelligence coverage.

Why seats are unlimited

Units already capture how much your team uses the platform. Charging per seat on top of that would double-count the same growth, so we don’t. Invite everyone.

What happens if you go over

Your traffic is never blocked, on any plan:
  1. We notify you at 80% and again at 100% of your plan’s monthly units. These are informational — nothing bills extra.
  2. On paid plans, requests keep serving and recording at full quality through the end of the period, no matter how far over you go.
  3. There’s a grace band of roughly 20% before a plan change is even considered, and only two consecutive periods over moves you up — one spiky month never bounces you.
  4. Plan moves take effect at the next billing period, with advance notice — and only upward. We never downgrade you automatically; dropping your plan is one click in Settings, anytime.
On the Free plan, routing never stops. Past the 10M-unit allowance, your requests keep flowing and stay in your history — they’re just locked from view until you upgrade (upgrading unlocks them retroactively) or your period resets and new traffic records again.

Add-on: managed dedicated deployment

Open-weight models in your own cloud account. We write the deployment, run it with you, monitor it, upgrade models, and keep parity proven — compute bills to your cloud directly, never through us. From 1,000setup,then1,000 setup, then 499/month per environment. Honestly: this only makes financial sense above roughly $2,000/month of model spend, or when data residency outweighs cost. Below that, serverless open-weights setup (Together, Fireworks, and similar) is included free on every plan.

Cancelling

Anytime, from Settings → Billing → Manage subscription. You keep your plan until the end of the period you paid for, then land back on Free. Your gates, keys, and analytics are untouched — and leaving Verlon entirely is a base-URL change, documented at verlon.ai/leave.