Skip to main content

The model

Everything that moves a request is free on every plan — routing across every provider, live model swapping, failover chains, spend caps, and bringing your own keys at zero markup. The subscription pays for the intelligence layer: the machinery that judges, compares, and proves your AI works on your real traffic. We never mark up tokens, never charge per seat, and never bill overages.

Plans at a glance

Rate limits exist on every plan as abuse prevention — they scale with your tier and are not something we sell.

Why seats are unlimited

Units already capture how much your team uses the platform. Charging per seat on top of that would double-count the same growth, so we don’t. Invite everyone.

What happens if you go over

Nothing dramatic, by design:
  1. We notify you at 80% and again at 100% of your plan’s monthly units. These are informational — nothing is blocked and nothing bills extra.
  2. Requests keep serving at full quality through the end of the period, no matter how far over you go.
  3. There’s a grace band of roughly 20% before a plan change is even considered, and only two consecutive periods over moves you up — one spiky month never bounces you.
  4. Plan changes take effect at the next billing period, with advance notice.
  5. It’s symmetric: if your volume drops for two periods, your plan drops with it. Never a one-way ratchet.
On the Free plan, sustained heavy use (two consecutive months well above the allowance) eventually pauses traffic on our keys until you upgrade — your own provider keys keep working the entire time.

Add-on: managed dedicated deployment

Open-weight models in your own cloud account. We write the deployment, run it with you, monitor it, upgrade models, and keep parity proven — compute bills to your cloud directly, never through us. From 1,000setup,then1,000 setup, then 499/month per environment. Honestly: this only makes financial sense above roughly $2,000/month of model spend, or when data residency outweighs cost. Below that, serverless open-weights setup (Together, Fireworks, and similar) is included free on every plan.

Cancelling

Anytime, from Settings → Billing → Manage subscription. You keep your plan until the end of the period you paid for, then land back on Free. Your gates, keys, and analytics are untouched — and leaving Verlon entirely is a base-URL change, documented at verlon.ai/leave.