Fragmented AI access
Teams integrate models, MCP servers and internal APIs independently, creating inconsistent access and duplicated infrastructure.
With Fabric: Connect every AI service through one governed gateway.
LLMs, MCP servers, APIs, and agents, routed, governed, and observed through a single endpoint deployed inside your own cloud.
50+ providers1,200+ modelsLLM, MCP, A2A and API supportFlexible deployment options
Fabric Gateway is an AI-native gateway built for agents, LLMs, MCP servers, and the systems they interact with.
Why Fabric Gateway
What breaks when AI adoption expands across teams—and how Fabric restores control.
Teams integrate models, MCP servers and internal APIs independently, creating inconsistent access and duplicated infrastructure.
With Fabric: Connect every AI service through one governed gateway.
Outages, latency spikes and rate limits can disrupt every application relying on a provider.
With Fabric: Use health-aware routing, retries, circuit breakers and automatic fallbacks.
Application-level guardrails and permissions drift as more teams and workloads are added.
With Fabric: Define policy centrally and enforce it across models, agents, tools and APIs.
Organizations struggle to understand which teams, applications and models are driving token usage and cost.
With Fabric: Attribute usage, set budgets and enforce token-aware limits in real time.
Security teams cannot reliably determine who initiated an action, which policies applied or what an agent accessed.
With Fabric: Capture identity, routing, policy and authorization outcomes for every governed request.
Fabric separates centralized configuration and governance from runtime traffic. The control plane manages policy and identity, while the data plane routes requests inside the selected deployment environment.
Your stack
Applications
SDK
AI agents
MCP
A2A
CLI tools
Sandboxed
execution
Developers
Any HTTP
SDK
Control plane
Postman Cloud
Registry
APIs · MCP · Agents
LLMs · CLIs
Policy & config manager
Versioning & syncing
Identity & RBAC
Org → team → agent
Admin portal
Analytics · Audit
UI & management
Data plane
Customer cloud
Protocol adapters
HTTP · gRPC
MCP · A2A
Plugin chain
Auth · Rate limit
PII · Transforms
Credential broker
OAuth · AES-256
Vault · Key management
Router
Radix tree
Weighted LB · Failover
OTel
Traces · Metrics · Logs
Providers
LLM providers
1,200+ models
across 50+ providers
MCP registry
Merged catalog
REST and gRPC APIs
Any HTTP backend
External & internal agents
SPIFFE trust
Platform checklist
Auto-failover
Circuit breaker & retry
Rate limits
Per-team token budgets
Observability
Cost, latency, audit logs
Secret isolation
Agents never see keys
Guardrails
PII, RBAC, audit trail
Cloud-native
AWS, GCP, Azure
Capabilities
Built for developers and designed for the enterprise.
Route models, MCP servers, agents and internal APIs through one endpoint, with consistent authentication, governance and observability across every connection.
Apply authentication, authorization, budgets and data controls centrally across teams, agents, models, tools and environments.
Agents use scoped Fabric tokens while the data plane retrieves and injects the correct upstream credential at request time.
Fabric understands models, tokens, streaming responses and tool calls, enabling token-aware limits, resilient routing and provider fallbacks.
Follow model, MCP, agent and API calls across one workflow, with identity, policy, latency, token and cost context.
Validate the initiating user and executing agent, restrict delegated permissions and deny requests when identity or policy cannot be verified.
Coverage
LLMs, MCP servers and your own services, registered once and governed the same way.
1,200+ models from 50+ providers through one OpenAI-compatible endpoint. Change the base URL; keep your SDK.
base_url = https://api.fabricgateway.ai/v1
// Point your SDK here, no other code changes.
1,200+ models across 50+ providers
Point a supported SDK at Fabric by changing the base URL.
Retries, fallbacks and circuit breakers across providers.
Apps never hold upstream provider keys.
Zero Trust
When work moves between agents and tools, Fabric carries only the identity and permissions each hop needs, and denies the request when it cannot verify either.
Check the initiating user and the executing agent on every request before anything runs.
Issue purpose-scoped access per hop instead of forwarding the original user token.
Allow only the tools and actions the agent is permitted to use.
If identity or policy cannot be verified, Fabric denies the request.
For the Enterprise
Choose the deployment model that fits your security review.
Run the data plane inside your own Kubernetes cluster or VM fleet.
Let Postman operate it for you with regional data planes.
Control plane managed by Postman; data plane stays in your cloud.
Connect your first model, MCP server or internal API and apply consistent governance from one gateway.