Skip to content

Envoy AI Gateway becomes Agent Router and joins the Agentic AI Foundation

Learn more

See Every Token. Manage Your AI Traffic. Operate Anywhere
at Scale.

Agent Router Enterprise is the AI gateway enterprises run in production: route to any model, attribute spend by team and agent, and set policies for your AI usage.
All enforced consistently everywhere you operate: any geo, any infra, any model.

Use Cases

Visibility

Cost attribution

Chargeback by team, agent and project — captured at the request, not reconstructed from three provider invoices six weeks later.

Agent traces

What an agent called, in what order, and where the latency went — across a whole chain, not one hop.

Audit evidence

Immutable audit logs, attributable by jurisdiction, with retention you can defend to a regulator months after the quarter closed.

Access

Access from the directory you already run

Okta, Entra, any OIDC provider. Directory groups are the unit of policy, so model access inherits the joiner-mover-leaver process your auditors already accept. Move teams on Monday, access moves on Monday.

Secure tool access MCP

A curated tool catalogue with per-profile authentication and filtering governs exactly what an agent can reach — so a sub-agent can never do more than the person it is acting for.

Control

Cap runaway agent costs

Token budgets and rate limits enforced inline stop a looping or misbehaving agent from burning spend before anyone notices it happened.

Policy and data-loss prevention

Runtime guardrails redact PII, filter prompts and block risky transactions before the damage reaches a live system.

Routing policy

High-volume work to the efficient model, the high-value business process to the most capable one — and a defined answer for when neither is available.

Enterprise Flexibility

Provider resilience

Automatic fallback and traffic splitting keep agents running through a provider outage — with no manual failover code in your application.

Regional and edge inference

Distributed gateways by region — or by zip code — cut latency and satisfy data-residency requirements in the same decision.

Regulated, high-scale operations

CVE-protected builds running in your own VPC, on-premises or air-gapped, supporting compliance at production scale.

Developer Productivity

Developer enablement

Claude Code, Claude Cowork, Codex, Cline, Cursor and any OpenAI-compatible harness work as they come. From install to building in minutes, with no agent rewritten and no key handed out.

Model evaluation

A built-in playground, A/B testing and traffic splitting pick the best cost-to-quality trade-off with evidence rather than opinion — before it becomes a standard.

Easy to Choose. Expensive to Change.

A Simple Starting Point

Every AI gateway will get you to your first request this afternoon. On day one, they’re close to interchangeable, and the comparison tables all say the same things — ours included.

A Long-Term Decision

A gateway sits in the path of every AI request your company makes. As your usage changes and infrastructure grows, replacing it becomes a costly decision you won’t want to make twice a year.

Your AI Infrastructure Grows in Two Directions

Your gateway needs to expand across regions and infrastructure while handling increasing traffic, concurrency, and complexity.

IT SPREADS OUT

More Places to Enforce

  • Latency wants a gateway near your agents. Every hop in a chain pays the trip.
  • Sovereignty wants one per jurisdiction, each with a different permitted-model list.
  • Availability wants one in the region you fail over to — your gateway's uptime is now your product's.
  • Private GPU capacity wants one close enough to each cluster to see how busy it is.

IT SCALES UP

More Load on Each One

  • Throughput turns a rate-limit counter into distributed accounting under streaming responses.
  • Concurrency turns a budget check into a hot path with a latency budget of its own.
  • Chain depth multiplies one user request into fifty model calls, each needing a decision.
  • Volume makes attribution a data problem long before it becomes a reporting problem.

A Unified Gateway Between Your Agents and Every Model