Compare your monthly costs
Compare input, output, and cache rates against your current usage. Share your approximate monthly spend so we can review your workload.
For teams
Lower inference costs for your product or business. Compare rates for the models you use, then tell us your volume and billing needs before you move traffic.
Compare input, output, and cache rates against your current usage. Share your approximate monthly spend so we can review your workload.
Tell us which models matter, how your traffic varies, and when you want to start. We onboard gradually as capacity becomes available.
Need company invoices or a purchase order reference? Include your billing requirements in the request. We’ll confirm available options before onboarding.
Tell us about procurement, data handling, and support requirements. Review the details with us before moving production traffic.
Use the same OpenAI-compatible API for a prototype and a production application. Check your models, features, and expected usage before changing traffic.
Access is by invitation for individuals and businesses. Send a request to join the onboarding list; we’ll email you when we can support your workload.
Explore models and detailed rates →Share your models, monthly spend, and billing requirements in one request.