Apache APISIX
No vendor lock-in
Open-source data plane, open plugin model, and portable configuration for APIs and AI workloads.
Apache APISIX AI Gateway
Use Apache APISIX as an open-source LLM gateway and proxy for model routing, load balancing, retries, fallback, token rate limiting, security, and observability.
One gateway, two workloads
Keep the routing, security, observability, and operations model your teams already use while adding LLM traffic.
Apache APISIX
Open-source data plane, open plugin model, and portable configuration for APIs and AI workloads.
100+
Authentication, traffic control, observability, serverless, and AI plugins.
Load balancing, fallback, token controls, RAG, prompt policies, moderation, and auditing are all available as open-source plugins.
Architecture
Built for production AI traffic
Route across OpenAI, DeepSeek, Claude, Mistral, Gemini, and other providers with health checks and weighted balancing.
Control token consumption by Route, Service, Consumer, Consumer Group, or custom attributes in standalone and cluster deployments.
Connect enterprise knowledge to model requests at the gateway layer for grounded, context-aware responses.
Track token usage through access logs and existing observability tools to control abuse and unexpected cost.
Use health checks, automatic retries, and fallback providers to keep AI applications available when an upstream model fails.
Apply prompt guards, decorators, templates, content moderation, logging, and auditing before traffic reaches a model.
Provider choice
Use OpenAI, DeepSeek, Claude, Mistral, Gemini, self-hosted models, or a mix of providers without changing the application-facing gateway.
See the AI Gateway capabilities