VarunSH

Pricing

Pay for tokens, not surprises

Every plan includes the full API, scoped keys, and usage analytics. Quotas are enforced server-side — no overage shocks.

Free

$0/mo

5M tokens / month

  • Access to eligible free models
  • API access
  • Basic usage analytics
  • API keys
  • Community support
Start free

Starter

$10/mo

30M tokens / day

  • Expanded model access
  • Higher rate limits
  • Usage analytics
  • Priority routing
  • More concurrent requests
Choose Starter
Most popular

Growth

$30/mo

100M tokens / day

  • Higher limits
  • Premium model access
  • Higher concurrency
  • Advanced analytics
  • Priority routing
Choose Growth

Scale

$50/mo

500M tokens / day

  • Maximum standard limits
  • Premium model access
  • High concurrency
  • Priority routing
  • Advanced analytics
  • Production workloads
Choose Scale

All quotas and limits are configurable and enforced server-side. Billing integration is provider-agnostic and enabled per environment.

Pricing questions

A unified API gateway that routes requests to your AI models through one OpenAI- and Anthropic-compatible endpoint. Point your existing SDK at our base URL and go.

Rarely. If you already use the OpenAI or Anthropic SDK, you only change the base URL and API key. The request and response shapes are compatible.

Usage is metered on actual tokens returned by the upstream model. Each plan includes a token allowance; pricing per model is transparent on the model's detail page.

No. By default we record only metadata needed for usage accounting — tokens, latency, status, model. Prompt and response bodies are never persisted.

When fallback routing is enabled for a model, requests automatically retry against the next healthy provider route in priority order, then return a normalized response.

Yes. The platform is provider-agnostic. An administrator adds a provider, configures its endpoint and credentials, maps a model, and it's live — no frontend changes.