AI Gateway
An API Evangelist landscape index of AI gateways — the LLM routers, prompt firewalls, model fallback proxies, cost-control planes, and policy engines that sit between applications and AI providers. AI gateways unify access across OpenAI, Anthropic, Google, AWS Bedrock, Azure OpenAI, and self-hosted models behind a common interface and apply caching, routing, guardrails, observability, rate limiting, budgets, RBAC, and audit controls. This index catalogs commercial SaaS gateways, open-source projects, API gateway AI plugins, and cloud-provider AI proxies, with a shared schema and vocabulary for describing model routes, fallbacks, guardrails, and budgets across vendors.
Limited machine-readable signal and partial portal coverage — documentation a human can read, but little a machine or agent can consume without scraping.
API Evangelist profiles AI Gateway the way a machine reads it — 108 machine-readable artifacts across 1 API, pulled from the provider's own public surface and indexed so a developer, an analyst, or an AI agent can evaluate it against every other provider on the network.
Every provider in the network is reduced to the same set of machine-readable artifacts — OpenAPI contracts, event specifications, GraphQL schemas, runnable collections, pricing and rate-limit signals, security posture, OAuth scopes, and the agent surfaces (MCP servers and skills) that let software drive the API on its own. We profile them because the interface is the part of a company you can actually inspect: it is a truer signal of what a provider does than any marketing page. From those artifacts we compute the Kin Score — AI Gateway scores 38.1/100 (thin), with a separate agent-readiness read of 20/100 (agent aware). The full breakdown is below, followed by every artifact we hold — each card links through to its machine-readable definition on apis.io.
Kin Score
This is the API Evangelist rating — a single, repeatable read computed from the artifacts on this page. Green fill is points earned; the red track is points possible, so every bar shows earned-versus-possible at a glance. Every facet and dimension name is a link: it opens that measurement's page on APIs.io, where the rating runs across the whole catalog — the exact checks that feed it, how every profiled provider distributes on it, and who is at the top of it.
Put this on your own site. The badge is drawn live from AI Gateway's current Kin Score — paste it once and it updates itself every time the score is recomputed. It follows your visitor's light or dark setting, and it links back here so anyone who sees it can read the full breakdown.
<!-- Kin Score · API Evangelist -->
<a href="https://providers.apievangelist.com/providers/ai-gateway/"
title="AI Gateway on API Evangelist — API profile and Kin Score">
<img src="https://apis.io/badge/ai-gateway.svg"
alt="AI Gateway Kin Score — API readiness rating by API Evangelist" width="150" height="150" loading="lazy">
</a>
[](https://providers.apievangelist.com/providers/ai-gateway/)
<!-- Kin Score · API Evangelist -->
<a href="https://providers.apievangelist.com/providers/ai-gateway/"
title="AI Gateway on API Evangelist — API profile and Kin Score">
<img src="https://apis.io/badge/ai-gateway/card.svg"
alt="AI Gateway Kin Score — API readiness rating by API Evangelist" width="340" height="120" loading="lazy">
</a>
More shapes, themes and sizes → · Score as JSON · How badges work
How we profile AI Gateway
Each block below is one kind of artifact we hold for AI Gateway. For each we say what it is and why it earns a place in the profile, then list every one we've indexed — capped at two rows, scroll within the panel for the rest.
APIs 38
Each API is captured as its own OpenAPI definition — every operation, parameter, and response. This is the single most useful machine-readable description of what an API does, and it's what lets us score, lint, mock, and generate against it without asking the provider for anything.
Individual APIs this provider publishes, each with its own machine-readable definition.
Portkey
Portkey is a production-grade AI gateway and control plane that fronts 1,600+ LLMs with unified routing, fallbacks, semantic caching, guardrails, cost attribution, and prompt ma...
OpenRouter
OpenRouter is a unified inference marketplace exposing 400+ models from 60+ providers behind one OpenAI-compatible API, with automatic provider fallback, pay-as-you-go credits, ...
LiteLLM
LiteLLM (BerriAI) is an open-source LLM gateway that exposes 100+ LLM providers — OpenAI, Anthropic, Azure, Bedrock, Gemini — through a single OpenAI-compatible API. The LiteLLM...
Helicone
Helicone is an open-source AI observability and routing platform centered on requests, sessions, prompts, datasets, rate limits, and alerts. Integrates with OpenAI, Anthropic, G...
Cloudflare AI Gateway
Cloudflare AI Gateway is an edge-deployed proxy that fronts AI providers — Workers AI, Anthropic, Google Gemini, OpenAI, Replicate, and more — with caching, rate limiting, analy...
Kong AI Gateway
The Kong AI Gateway is delivered as the AI Proxy plugin for Kong Gateway, transforming and proxying requests across 16+ providers including OpenAI, Azure OpenAI, Anthropic, Amaz...
Apache APISIX AI Proxy
The Apache APISIX ai-proxy plugin streamlines integration with LLMs by converting plugin settings into the appropriate request format for OpenAI, DeepSeek, Azure OpenAI, Anthrop...
Tetrate Agent Router Service
Tetrate Agent Router Service is an Envoy AI Gateway-as-a-service from the creators of Envoy, providing an approved LLM catalog, unified model access, automatic fallback, cost ma...
NVIDIA NIM
NVIDIA NIM is a set of inference microservices for streamlined AI model deployment, prebuilt and optimized for low-latency, high-throughput inference on NVIDIA-accelerated infra...
Traefik AI Gateway
Traefik AI Gateway is an enterprise, self-hosted, Kubernetes-native AI gateway with safety and governance (NVIDIA Safety NIMs, jailbreak detection, content filtering across 22+ ...
Together AI
Together AI is a full-stack AI Native Cloud for inference, fine-tuning, and GPU clusters powered by research, exposing serverless inference, batch processing, dedicated model an...
Anyscale
Anyscale is the production-scale AI platform built on Ray by the creators of Ray, supporting LLM inference and other data-intensive AI workloads across distributed GPU clusters....
LangDB
LangDB is an enterprise AI gateway for routing and governing LLM traffic across providers, with observability, cost tracking, and policy enforcement. Public homepage was unreach...
Envoy AI Gateway
Envoy AI Gateway is an open-source extension to Envoy Proxy and Envoy Gateway, providing a Kubernetes-native AI traffic plane for routing, governing, and observing LLM calls acr...
Gentrace
Gentrace was an AI evaluation and observability product; the company has shut down and its codebase is now MIT-licensed open source on GitHub. Included here for historical compl...
AI Gateway Analytics API
The Analytics API from AI Gateway — 2 operation(s) for analytics.
AI Gateway APIKeys API
The APIKeys API from AI Gateway — 1 operation(s) for apikeys.
AI Gateway Assistants API
The Assistants API from AI Gateway — 1 operation(s) for assistants.
AI Gateway Audio API
The Audio API from AI Gateway — 3 operation(s) for audio.
AI Gateway Batches API
The Batches API from AI Gateway — 2 operation(s) for batches.
AI Gateway Chat API
The Chat API from AI Gateway — 1 operation(s) for chat.
AI Gateway Completions API
The Completions API from AI Gateway — 1 operation(s) for completions.
AI Gateway Configs API
The Configs API from AI Gateway — 1 operation(s) for configs.
AI Gateway Embeddings API
The Embeddings API from AI Gateway — 2 operation(s) for embeddings.
AI Gateway Feedback API
The Feedback API from AI Gateway — 1 operation(s) for feedback.
AI Gateway Files API
The Files API from AI Gateway — 2 operation(s) for files.
AI Gateway FineTuning API
The FineTuning API from AI Gateway — 2 operation(s) for finetuning.
AI Gateway Guardrails API
The Guardrails API from AI Gateway — 1 operation(s) for guardrails.
AI Gateway Images API
The Images API from AI Gateway — 1 operation(s) for images.
AI Gateway Integrations API
The Integrations API from AI Gateway — 2 operation(s) for integrations.
AI Gateway Logs API
The Logs API from AI Gateway — 2 operation(s) for logs.
AI Gateway MCP API
The MCP API from AI Gateway — 2 operation(s) for mcp.
AI Gateway Policies API
The Policies API from AI Gateway — 2 operation(s) for policies.
AI Gateway Prompts API
The Prompts API from AI Gateway — 6 operation(s) for prompts.
AI Gateway Responses API
The Responses API from AI Gateway — 1 operation(s) for responses.
AI Gateway Threads API
The Threads API from AI Gateway — 3 operation(s) for threads.
AI Gateway VirtualKeys API
The VirtualKeys API from AI Gateway — 1 operation(s) for virtualkeys.
AI Gateway Workspaces API
The Workspaces API from AI Gateway — 4 operation(s) for workspaces.
Scroll within the panel for all 38 ·
Open Collections 25
Open, tool-agnostic collections carry the same runnable value as Postman without locking you to one client — the portable, forkable form of the same exercise.
Open, tool-agnostic API collections (OpenAPI-derived and Bruno).
API Collection
OPEN COLLECTIONPortkey AI Gateway Analytics API
OPEN COLLECTIONPortkey AI Gateway Analytics APIKeys API
OPEN COLLECTIONPortkey AI Gateway Analytics Assistants API
OPEN COLLECTIONPortkey AI Gateway Analytics Audio API
OPEN COLLECTIONPortkey AI Gateway Analytics Batches API
OPEN COLLECTIONPortkey AI Gateway Analytics Chat API
OPEN COLLECTIONPortkey AI Gateway Analytics Completions API
OPEN COLLECTIONPortkey AI Gateway Analytics Configs API
OPEN COLLECTIONPortkey AI Gateway Analytics Embeddings API
OPEN COLLECTIONPortkey AI Gateway Analytics Feedback API
OPEN COLLECTIONPortkey AI Gateway Analytics Files API
OPEN COLLECTIONPortkey AI Gateway Analytics FineTuning API
OPEN COLLECTIONPortkey AI Gateway Analytics Guardrails API
OPEN COLLECTIONPortkey AI Gateway Analytics Images API
OPEN COLLECTIONPortkey AI Gateway Analytics Integrations API
OPEN COLLECTIONPortkey AI Gateway Analytics Logs API
OPEN COLLECTIONPortkey AI Gateway Analytics MCP API
OPEN COLLECTIONPortkey AI Gateway Analytics Policies API
OPEN COLLECTIONPortkey AI Gateway Analytics Prompts API
OPEN COLLECTIONPortkey AI Gateway Analytics Responses API
OPEN COLLECTIONPortkey AI Gateway Analytics Threads API
OPEN COLLECTIONPortkey AI Gateway Analytics VirtualKeys API
OPEN COLLECTIONPortkey AI Gateway Analytics Workspaces API
OPEN COLLECTIONPortkey AI Gateway API
OPEN COLLECTIONScroll within the panel for all 25 ·
Features 13
The notable capabilities this provider advertises, captured as structured features so they can be searched and compared instead of read one landing page at a time.
Notable capabilities this provider offers.
Provider Abstraction
A unified, typically OpenAI-compatible API surface that lets clients call any supported LLM provider without provider-specific SDK juggling.
Model Routing
Route requests to the right model and provider based on alias, header, request content, identity, time-of-day, cost, or latency.
Fallback and Failover
Automatically retry failed requests against backup providers or models when a primary upstream is degraded, rate-limited, or down.
Load Balancing and Fanout
Distribute traffic across multiple providers or replicas using weighted, priority-based, or RPM/TPM-aware load balancing.
Response Caching
Exact-match and semantic caching of model responses to cut latency and provider spend; some gateways claim 40-70 percent cost savings.
Cost Controls and Budgets
Per-user, per-team, per-key, per-project budgets, spend tracking, and hard or soft caps on token consumption.
Rate Limiting and Quotas
RPM, TPM, concurrency, and per-key quotas enforced at the gateway, decoupled from each upstream provider's limits.
Guardrails and Prompt Firewall
Prompt injection detection, jailbreak filtering, content moderation, PII redaction, and topic control applied to requests and responses.
Observability
Request, response, token, cost, latency, error, and trace data exported via OpenTelemetry, Langfuse, Phoenix, Langsmith, or built-in dashboards.
Authentication and RBAC
Virtual keys, JWT, OAuth2, SSO, and role-based access control over which clients can use which models with which budgets.
BYOK and Secret Management
Bring-your-own provider API keys, with the gateway holding and injecting them so clients never see upstream credentials.
Multi-Tenant Governance
Per-tenant isolation of keys, budgets, logs, and policies for platform teams serving multiple internal product teams.
MCP Federation
Some AI gateways also front Model Context Protocol servers, aggregating tools and exposing a single MCP endpoint to agents.
Scroll within the panel for all 13 ·
Semantic Vocabularies 1
JSON-LD contexts give the data shared meaning across APIs. We profile them because semantics are what let a machine reconcile 'customer' here with 'customer' somewhere else.
JSON-LD contexts and semantic vocabularies used across these APIs.
Ai Gateway Context
JSON-LDSpectral Rules 1
Governance rulesets we run against this provider's specs — the automated checks behind parts of the score. Profiling them makes the quality bar explicit and re-runnable, not a matter of opinion.
AI Gateway API Rules
SPECTRALJSON Schema 3
Standalone JSON Schema definitions describe the data models behind the API. We profile them so the shapes are validatable on their own — useful long after a single request is forgotten.
Standalone JSON Schema definitions for this provider's data models.
JSON Structure 3
JSON Structure captures the data shapes in a form built for tooling — a complement to JSON Schema that keeps the model machine-legible.
JSON Structure definitions describing this provider's data shapes.
Ai Gateway Policy Structure
JSON STRUCTUREAi Gateway Provider Structure
JSON STRUCTUREAi Gateway Route Structure
JSON STRUCTUREExamples 6
Real request and response payloads are what turn a spec from abstract into obvious — and they're one of the twelve things an agent needs to call an API correctly on the first try.
Example request and response payloads for these APIs.
Ai Gateway Provider Example
EXAMPLEAi Gateway Route Example
EXAMPLESecurity Posture 2
Authentication, domain security, vulnerability disclosure, and trust-center signals — the evidence that a provider takes security seriously enough to document it. We profile it because you can't govern what you can't see.
Authentication, domain security, vulnerability disclosure, and trust-center signals.
Agentic Access 1
An x-agentic-access contract marks which operations are safe for an agent to run on its own and which need a human in the loop. It is the difference between an API an agent can use and one it can use safely.
Recommended x-agentic-access execution contracts for AI agents.
Use Cases 6
What developers actually build with this provider — captured so the catalogue answers 'what is this for', not just 'what does this expose'.
What developers build with this provider.
Provider-Agnostic LLM Access
Front many LLM providers behind one API so application teams can switch models without changing client code.
Cost Containment for AI
Apply caching, routing to cheaper models, and per-team budgets to keep generative-AI spend predictable.
Reliability and Failover
Survive single-provider outages by automatically failing over to backup models when the primary degrades.
Centralized AI Governance
Enforce content, PII, and policy controls in one place for every AI request leaving the organization.
Observability and FinOps
Attribute cost and latency to teams, projects, and users; expose token-level metrics to FinOps and SRE.
Multi-Tenant AI Platforms
Build internal AI platforms where each product team gets its own virtual keys, budgets, and logs.
Integrations 9
Pre-built integrations with other platforms tell you where this provider already fits in a stack.
Pre-built integrations with other platforms and tools.
OpenAI
Front OpenAI's GPT, embeddings, and image models behind the gateway with virtual keys and budgets.
Anthropic
Route Claude requests through the gateway for fallback, caching, and central observability.
Google Gemini and Vertex AI
Proxy Google Gemini and Vertex AI calls with OpenAI-format translation where supported.
AWS Bedrock
Bridge OpenAI-format clients to Bedrock-hosted Anthropic, Mistral, Cohere, Meta, and Amazon models.
Azure OpenAI
Route to Azure-hosted OpenAI deployments with per-region failover and key rotation.
Ollama and vLLM
Front self-hosted Ollama and vLLM inference servers for hybrid cloud and on-prem inference.
OpenTelemetry
Export request, token, cost, and trace data to any OTel-compatible observability backend.
Langfuse and Phoenix
Stream prompts, completions, and evaluations to Langfuse and Arize Phoenix for prompt and model analytics.
Model Context Protocol
Some AI gateways federate MCP servers alongside LLM routes, exposing a unified agent endpoint.
Scroll within the panel for all 9 ·
Resources
Every other property we hold for AI Gateway — documentation, portals, status pages, policies, and corporate surface — grouped by the job it does, following the integrator's arc from getting started to running in production.
Get Started 1
Portal, sign-up, and the first successful call
Documentation 4
Reference material describing how the API behaves
Agent Surfaces 1
MCP servers, agent skills, and machine-readable catalogs
Design & Contract 5
Pagination, idempotency, versioning, errors, and events
Build 2
SDKs, sample code, and the tooling you integrate with
Access & Security 3
Authentication, authorization, and security posture
Operate 2
Status, limits, changes, and where to get help
Commercial 1
Pricing, plans, and the legal terms of use
Company 1
The organization behind the API
← All providers · Data indexed from github.com/api-evangelist/ai-gateway · machine-readable index on apis.io
This is an independent, third-party profile of AI Gateway, published by API Evangelist. We do not operate, host, resell, or support these APIs, and we are not affiliated with or endorsed by the company unless stated above. Everything here is built from publicly available information — the company's own site, developer portal, documentation, public repositories, and the specifications it publishes for public use. Nothing is obtained by breaching a system, defeating an access control, or using credentials.
The Kin Score and Agent Readiness rating are independently calculated assessments of a company's public API artifacts, scored against a published rubric. They are not certifications, endorsements, security assessments, or audits.
Corrections, re-scores, and removal are free — no partnership or purchase required, and you do not need to justify the request. A removed company is recorded as unrated, never scored zero for having asked. Acknowledgement within one business day; removal within two.
info@apievangelist.com
·
Read the full data-sourcing policy →
On a security or compliance team? Put security in the subject line and
you will get a person, not a form — we will tell you exactly which public URLs this profile was built from.