Business & Pricing
Plans, who this is actually built for, and how to reach a human.
Who it's for
Engineering teams
Shipping an LLM-backed product who want one integration, real provider flexibility, and visibility into cost and latency without building it themselves.
Platform teams
Centralizing AI access across multiple internal teams, with per-org isolation, RBAC, and an audit trail that satisfies a security review.
Cost-conscious teams
Using routing, failover, and caching to cut spend and avoid getting stuck on one provider's roadmap.
Plans
| Plan | Price | Fit |
|---|---|---|
| Free | $0 | Evaluate the gateway, routing, and the BYOK model at small scale. |
| Business | $149/mo | Production teams that need higher entitlements, policy control, and full dashboard depth. |
| Enterprise | Custom | Custom entitlements and direct support. If SSO or SOC 2 is a requirement, talk to us directly — see Compliance posture. |
How metering actually works today
Plan copy mentions usage-based inference; in practice, the usage tracking feeding the dashboard right now is request/estimate-based rather than full per-token metering. Feature gating always runs off entitlements, never the plan name directly. Roadmap has the current status of full metering.
Why teams actually save money on this
- Routing to the right model per request instead of one default model for everything.
- Caching skips paying for near-duplicate prompts twice.
- Failover keeps a request succeeding instead of getting retried by hand at the application layer.
- One integration across 18 providers means you're not maintaining N provider SDKs yourself.
Contact
- Sales / Enterprise
- kiran@inferect.online
- Support
- find@inferect.online
- Website
- https://inferect.online
Related