Pricing

One backend, four plans. Bring your own provider keys; each plan carries a monthly request and token allowance. Pick the tier that matches your traffic.

Free
$0

Evaluate and prototype.

  • Requests / mo10K
  • Projects1
  • Log retention7 days
  • Team seats1
  • Tokens1M / mo
  • Generative mediaNot included
  • SupportCommunity
Hobby
$24.99/ mo

Ship a real app, solo.

  • Requests / mo500K
  • Projects5
  • Log retention60 days
  • Team seats1
  • Tokens50M / mo
  • Generative mediaImage · audio · video (your keys)
  • SupportEmail
ProPopular
$99/ mo

Scale with traffic and a team.

  • Requests / mo5M
  • Projects25
  • Log retention1 year
  • Team seats5
  • Tokens500M / mo
  • Generative mediaImage · audio · video (your keys)
  • SupportPriority
Enterprise
Custom

High volume, SSO, and an SLA.

  • Requests / moCustom
  • ProjectsCustom
  • Log retentionCustom
  • Team seatsCustom
  • TokensCustom
  • Generative mediaImage · audio · video (your keys)
  • SupportSLA · SSO · dedicated

Every plan is bring-your-own-keys: your provider bills you for tokens, never us — the monthly token allowance on each card is this backend's own ceiling, not a bill. Generative media routes to your own media keys under a per-project spend budget. It isn't a charge we collect.

Questions

What counts as a request?
One inference call to the relay: a single turn of your agent. Tool calls inside a turn don't each count; a request is metered as one usage record.
Why bring your own keys?
You connect your own Anthropic, OpenAI, Gemini, and media-provider keys. Your provider bills you directly for tokens, so we never mark them up. You pay only the provider's cost, and our price is for the backend.
Is there a token limit?
Yes. Each plan carries a monthly token allowance, shown on its card, and a call is refused once the month's usage reaches it. Your provider still bills you for the tokens themselves; the allowance is the backend's own ceiling. Per-request ceilings you set on your agent profiles protect you from a runaway loop inside a single call.
What is log retention?
How long your usage and audit history stays queryable in the dashboard before it's purged. Each plan card shows its own window.
How does generative media work?
Image, audio, and video generation route to your own provider keys (OpenAI, ElevenLabs, Gemini, Veo…) under a per-project spend budget you set. Available from Hobby up.
What happens above Pro?
Talk to us. Enterprise covers custom request volume, SSO, an SLA, and dedicated support.
Can I change plans?
Yes. Upgrade or downgrade any time; the new limits apply from the change.