Skip to content

InferencePortkey

Portkey

AI gateway with observability, guardrails, and governance.

Pricing
FREEMIUM
Source
Open core
Hosting
Hybrid
Platforms
APIWeb
Models
Multi-model
Verified
Jun 8, 2026

A production AI gateway that gives apps and agents unified access to 1,600+ LLMs across providers behind a single API, with built-in observability, prompt management, guardrails, and governance. Portkey adds routing, caching, fallbacks, cost limits, PII redaction, RBAC, and an MCP gateway. Its core gateway is open-source; run it self-hosted/hybrid or use the managed cloud, which offers a free tier.

Capabilities 6

What it actually does — grouped by capability family.

  • MCP gateway / registry (secondary capability)
  • LLM gateway / routing (primary capability)
  • Multi-model access (secondary capability)
  • LLM observability (primary capability)
  • Prompt management (secondary capability)
  • Guardrails (secondary capability)

Pros & cons

  • One API across many providers
  • Routing, caching, and fallbacks
  • Cost limits and PII redaction
  • Guardrails, RBAC, and MCP gateway
  • Acquired by Palo Alto Networks (closed 2025)
  • Future roadmap may shift to Prisma AIRS
  • Advanced governance gated to paid tiers

Tags

View all Inference
  • View LiteLLM details
    InferenceFREEMIUMOpen core

    LiteLLM

    BerriAI

    AI gateway: call many LLMs through one OpenAI-format interface.

    Open-source Python SDK and proxy server (AI gateway) that exposes 100+ LLM providers through a single OpenAI-compatible API, with cost tracking, load balancing, fallbacks, caching, and guardrails. Self-host the proxy or use the managed cloud; a paid Enterprise tier adds SSO, audit logs, and support.

    Load balancing and guardrails built in
    Proxy adds an extra hop
    • gateway
    • proxy
    • routing
    • open-source
    • +1
  • View OpenRouter details
    InferenceFREEMIUM

    OpenRouter

    OpenRouter

    One OpenAI-compatible API in front of models from every provider.

    A unified gateway that routes a single endpoint and API key to models from Anthropic, OpenAI, Google, Meta, DeepSeek, xAI, and more — swap models by changing one parameter, with automatic fallbacks and one consolidated bill. Pass-through token pricing plus dozens of free models.

    Swap models by changing one parameter
    Adds a routing hop vs direct provider
    • gateway
    • routing
    • multi-model
    • fallbacks
  • View Helicone details
    ObservabilityFREEMIUMOpen core

    Helicone

    Helicone

    Drop-in LLM proxy with logging, caching, and cost tracking.

    One-line integration — change your OpenAI/Anthropic base URL and get a dashboard with every prompt, response, latency, and dollar tracked. Adds caching and rate-limit handling without code changes.

    No SDK or code changes to integrate
    Request/response focused, not span-based
    • proxy
    • logging
    • caching
    • cost-tracking
  • View Langfuse details
    ObservabilityFREEMIUMOpen core

    Langfuse

    Langfuse

    Open-source LLM observability. Self-hostable, OpenTelemetry-native.

    Tracing, evals, prompt management, and dataset tooling for LLM apps — self-host on your own infra or use Langfuse Cloud. The open-source default when you want full ownership of your observability stack.

    Own your observability data
    Self-host infra cost at scale
    • open-source
    • tracing
    • evals
    • self-hosted