Skip to main content
It takes 2 mins to integrate and with that, it starts monitoring all of your LLM requests and makes your app resilient, secure, performant, and more accurate at the same time.

Integrate in 3 Lines of Code

FAQs

The AI Gateway is hosted on edge workers throughout the world, ensuring minimal latency. Our benchmarks estimate a total latency addition between 20-40ms compared to direct API calls. This slight increase is often offset by the benefits of our caching and routing optimizations.
All data is encrypted in transit and at rest using industry-standard AES-256 encryption, and the gateway follows best practices for service security, data storage, and retrieval.
The gateway is built on scalable infrastructure and can handle millions of requests per minute with very high concurrency. Its edge architecture and scaling capabilities accommodate sudden spikes in traffic without performance degradation.
No explicit timeout is imposed on requests. While the gateway does not time out requests on our end, we recommend implementing client-side timeouts appropriate for your use case to handle potential network issues or upstream API delays.
Yes! We support SSO with any custom OIDC provider.
Last modified on September 15, 2026