Skip to content
Inferect
DocsBusiness

Business & Pricing

Plans, who this is actually built for, and how to reach a human.

5 min read

Who it's for

Engineering teams

Shipping an LLM-backed product who want one integration, real provider flexibility, and visibility into cost and latency without building it themselves.

Platform teams

Centralizing AI access across multiple internal teams, with per-org isolation, RBAC, and an audit trail that satisfies a security review.

Cost-conscious teams

Using routing, failover, and caching to cut spend and avoid getting stuck on one provider's roadmap.

Plans

PlanPriceFit
Free$0Evaluate the gateway, routing, and the BYOK model at small scale.
Business$149/moProduction teams that need higher entitlements, policy control, and full dashboard depth.
EnterpriseCustomCustom entitlements and direct support. If SSO or SOC 2 is a requirement, talk to us directly — see Compliance posture.

How metering actually works today

Plan copy mentions usage-based inference; in practice, the usage tracking feeding the dashboard right now is request/estimate-based rather than full per-token metering. Feature gating always runs off entitlements, never the plan name directly. Roadmap has the current status of full metering.

Why teams actually save money on this

  • Routing to the right model per request instead of one default model for everything.
  • Caching skips paying for near-duplicate prompts twice.
  • Failover keeps a request succeeding instead of getting retried by hand at the application layer.
  • One integration across 18 providers means you're not maintaining N provider SDKs yourself.

Contact

Sales / Enterprise
kiran@inferect.online
Support
find@inferect.online
Website
https://inferect.online