Skip to main content

Pricing

Usage-based and simple. Pay only for what you use, with no subscription. Free during beta while we finish billing.

Free during beta, no charges today

You are not billed during beta and no credit card is required. The rates below are what applies when we turn billing on, and we will give existing users notice first.

Free

$0to start

Try smart routing and inference with free credit. No card required.

  • $5 inference credit on signup
  • 5,000 free routing decisions every month
  • Smart routing and custom model pools
  • OpenAI-compatible API
  • Model catalog, comparison, and playground
  • Organizations, projects, and role-based access
  • Community support
Start free
Most Popular

Pay-as-you-go

$1per 1,000 routed requests

Only pay for what you use, on two simple meters. No subscription.

Everything in Free, plus

  • $0.001 per routing decision beyond the free 5,000
  • Managed inference at provider cost plus a flat 6%
  • Pinned models are never charged a routing fee
  • No per-request minimum charge
  • $10 minimum top-up, cancel any time by not topping up
  • Higher rate limits
  • Usage dashboard with spend caps and budgets
Get started

Enterprise

Customlet's talk

For teams adopting AI at scale, with the controls and support you need.

Everything in Pay-as-you-go, plus

  • Volume and committed-use pricing
  • Dedicated capacity and priority routing
  • SSO and an uptime SLA
  • Custom model onboarding
  • Audit logs and compliance support
  • Dedicated support and onboarding
Contact sales

Two usage meters, billed separately: smart routing ($1 per 1,000 routed requests, first 5,000 each month free) and managed inference (provider token price plus a flat 6%). You are only charged the routing fee when routing chooses the model for you; pinning a model skips it.

Compare plans

Feature comparison across the Free, Pay-as-you-go, and Enterprise plans
FeaturesFreePay-as-you-goEnterprise
Smart routing
Automatic model selection
Free routing decisions5,000 / mo5,000 / moCustom
Routing fee beyond the free allowance$0.001 / decisionVolume pricing
Custom model pools
Eligibility presets
Automatic fallback and retries
Routing decision audit trail
Inference
OpenAI-compatible endpoint
Managed inference priceAt cost + 6%At cost + 6%Negotiated
Streaming responses
$5 signup creditCustom
Minimum top-up$10Invoicing
Platform
Model catalog and comparison
Playground
Organizations and projects
Role-based access control
Scoped API keys
Usage dashboard
Spend caps and budgets
GPU capacity planner
Scale and support
Rate limitsStandardHigherCustom
SupportCommunityCommunityDedicated
SSO
Uptime SLA
Dedicated capacity
Custom model onboarding
Audit logs and compliance

Standard rate limits apply to keep the platform stable. Enterprise features such as SSO, an SLA, audit and compliance, and dedicated capacity are available on request. Contact us for higher limits or a custom plan.

Frequently asked questions

Your AI stack shouldn't stand still.

Every month new models become cheaper, faster, and more capable. Inferbase ensures your application automatically benefits without changing a single API call.