Skip to main content

AI Gateways & LLM Routing — Market

Updated 6/19/2026

Verified claims and product-axis read for AI Gateways & LLM Routing. Every fact below is sourced; every product judgment traces back to underlying signals.


Verified facts

  • Portkey, LiteLLM, and Helicone all expose Prometheus-compatible metrics endpoints. (other)
  • OpenRouter's terms allow models to be deprecated with 14-day notice to API consumers. (other)
  • OpenRouter natively exposes provider-side privacy policies (training-data opt-out) per model. (other)
  • Helicone introduced cost alerts via Slack/webhook in 2024. _(historical_event)_
  • OpenRouter offers BYOK (bring your own key) at 0% markup, monetizing only on prepaid credits. (other)
  • LiteLLM project maintainer count grew to 12+ regular committers in 2025. _(historical_event)_
  • LiteLLM hosts router benchmarks showing <5ms p99 added latency on small-payload requests. _(technical_spec)_
  • Portkey integrates with AWS Bedrock, Azure OpenAI, and GCP Vertex out of the box. (other)
  • Portkey claims 99.99% uptime SLA on its Enterprise tier with regional failover. (other)
  • LiteLLM Proxy publishes weekly Docker pulls exceeding 500K on Docker Hub. (other)

Top products (engine read)

OpenAI-compatible Multi-Provider LLM Gateway — gateway / proxy

Opportunity: Devs want one drop-in OpenAI-compatible endpoint that abstracts provider sprawl, handles failover, and removes provider lock-in (fact 4f1adaf4: Portkey 250+ LLMs; fact 495572d5: LiteLLM Proxy 100+ providers).

This is the spine of the ai-gateways category. Portkey (7k+ stars per fact 6c959edf), LiteLLM, and Helicone are all hiring backend/platform engineers around it; community posts show indie 'AISBF' and 'AI Cost Firewall' clones validating product-market pull. The category is crowded — defensibility now shifts to ecosystem (caching, guardrails, observability) rather than the proxy itself.

LLM Guardrails / Agent Security Proxy — policy-enforcement gateway

Opportunity: Agent runtimes need an in-band judge to block prompt injection / data exfil / unsafe tool calls — Portkey ships 40+ pre-built deterministic & LLM-based checks (fact 6a2a6008-c864-48b8-b687-9d49125b6013); Brex open-sourced CrabTrap to fill the gap.

Strong tailwind: agentic workloads in prod are forcing gateways to grow a security plane. Helicone is hiring a 'Staff AI Agentic Security Engineer' and OpenRouter a 'Trust and Safety Lead' — every gateway is now also a policy gateway. Big-co (Brex) building in-house signals an enterprise willingness to pay.

Semantic Caching for LLM Gateway — cache layer

Opportunity: Cached requests are where gateways monetize latency + cost; Portkey advertises sub-1ms p50 added latency on cached calls (fact fd1ed84d-7563-4781-96e8-78ed2551fa44).

Semantic caching is moving from a Portkey feature to a standalone wedge — indie 'AI Cost Firewall' and AISBF lead with it as the headline. Hot, but commoditizing fast; whoever owns embeddings-aware cache + invalidation wins.


See the Products and Strategy modules for the full product list and forward-looking judgment.

Get this data as JSONLast updated: Jun 19, 2026