Pricing
Usage-based and simple. Pay only for what you use, with no subscription. Free during beta while we finish billing.
Free during beta, no charges today
You are not billed during beta and no credit card is required. The rates below are what applies when we turn billing on, and we will give existing users notice first.
Free
Try smart routing and inference with free credit. No card required.
- $5 inference credit on signup
- 5,000 free routing decisions every month
- Smart routing and custom model pools
- OpenAI-compatible API
- Model catalog, comparison, and playground
- Organizations, projects, and role-based access
- Community support
Pay-as-you-go
Only pay for what you use, on two simple meters. No subscription.
Everything in Free, plus
- $0.001 per routing decision beyond the free 5,000
- Managed inference at provider cost plus a flat 6%
- Pinned models are never charged a routing fee
- No per-request minimum charge
- $10 minimum top-up, cancel any time by not topping up
- Higher rate limits
- Usage dashboard with spend caps and budgets
Enterprise
For teams adopting AI at scale, with the controls and support you need.
Everything in Pay-as-you-go, plus
- Volume and committed-use pricing
- Dedicated capacity and priority routing
- SSO and an uptime SLA
- Custom model onboarding
- Audit logs and compliance support
- Dedicated support and onboarding
Two usage meters, billed separately: smart routing ($1 per 1,000 routed requests, first 5,000 each month free) and managed inference (provider token price plus a flat 6%). You are only charged the routing fee when routing chooses the model for you; pinning a model skips it.
Compare plans
| Features | Free | Pay-as-you-go | Enterprise |
|---|---|---|---|
| Smart routing | |||
| Automatic model selection | |||
| Free routing decisions | 5,000 / mo | 5,000 / mo | Custom |
| Routing fee beyond the free allowance | $0.001 / decision | Volume pricing | |
| Custom model pools | |||
| Eligibility presets | |||
| Automatic fallback and retries | |||
| Routing decision audit trail | |||
| Inference | |||
| OpenAI-compatible endpoint | |||
| Managed inference price | At cost + 6% | At cost + 6% | Negotiated |
| Streaming responses | |||
| $5 signup credit | Custom | ||
| Minimum top-up | $10 | Invoicing | |
| Platform | |||
| Model catalog and comparison | |||
| Playground | |||
| Organizations and projects | |||
| Role-based access control | |||
| Scoped API keys | |||
| Usage dashboard | |||
| Spend caps and budgets | |||
| GPU capacity planner | |||
| Scale and support | |||
| Rate limits | Standard | Higher | Custom |
| Support | Community | Community | Dedicated |
| SSO | |||
| Uptime SLA | |||
| Dedicated capacity | |||
| Custom model onboarding | |||
| Audit logs and compliance | |||
Standard rate limits apply to keep the platform stable. Enterprise features such as SSO, an SLA, audit and compliance, and dedicated capacity are available on request. Contact us for higher limits or a custom plan.
Frequently asked questions
Two usage meters, no subscription. Smart routing is $1 per 1,000 routing decisions, with the first 5,000 each month free and pinned models never charged a routing fee. Managed inference is the model provider token price plus a flat 6%. New accounts start with $5 of free inference credit after email verification.
No. Everything is free during beta and no payment details are required. The rates on this page apply only once we turn billing on, and existing users will get advance notice first.
No. Inferbase is pay-as-you-go. You add credit when you want it and pay only for what you use. There are no seats and no recurring platform fee.
$10. New accounts also receive $5 of free inference credit after verifying their email, so you can try routing and inference before adding anything.
Team and enterprise features are available on request: dedicated capacity, SSO, an SLA, audit and compliance, and custom onboarding. Contact us and we will work out the details.
Your AI stack shouldn't stand still.
Every month new models become cheaper, faster, and more capable. Inferbase ensures your application automatically benefits without changing a single API call.