Sovereign AI Gateway
Every AI call your company makes, governed
One gateway for 500+ models. Every tool call runs through the same policies, guardrails, and audit trail as your model calls. Start free and roll out across your company.
Ready to build?
Rolling out company-wide?
Providers
Models
Uptime

Trusted by AI teams running in regulated industries

We chose to work with Orq.ai to replace our internal setup with a production-ready AI Gateway that meets our governance, scalability, and cost-monitoring requirements.

Benjamin Kleppe,
GenAI Lead at bunq



The problem
How many agents is your company running right now?
What can they touch? What do they cost? Who can stop them?
How many agents run right now?
With Orq.ai every agent, model, and request routes through one gateway. So you always know what's running.
What can they touch?
Access control at the tool level. See and limit exactly what every agent can reach.
What do they cost right now?
Real-time spend attribution by team, project, or agent. No more bills that surprise you at month end.
Who can stop them in the next five minutes?
Kill switches and audit trails on every agent.
AI Governance
Govern every agent, model, and request
Every AI request passes through Orq.ai before it reaches a model or a tool. Security, compliance, cost, and operational policies are enforced right there, in real time.
Sees everything in real time
Every prompt, model call, tool call, and token, tracked as it happens.
Intervenes before execution
Block, reroute, redact PII, or enforce a budget cap, in real time, before a response is returned.
Central control, decentralized delivery
The central team sets the standards. Domain teams self-serve within them. No bottleneck, no shadow AI.
The Gateway
An AI gateway for every model, every request, every cost
Most AI gateways just pass keys. Most observability tools just watch. Most teams bolt on policies afterward. Orq.ai does all three, at the gateway level.
Model access
One API key, 500+ models across 30+ providers.
One decision point
Route traffic and read the traces in one dashboard, with cost attribution by user or team.
Reliability
Stay up when a provider goes down. Fallbacks, retries, caching, and load balancing.
Framework agnostic
Works with OpenAI Agents, Vercel AI, CrewAI, LangGraph, and any OpenTelemetry stack.
Policies that enforce themselves
Routing rules, PII redaction, and automatic failovers, applied to every request that matches.

Attribution by team, project, customer, or agent
Hard budget caps
Set limits at the org, team, project, model, or agent level. When a cap hits, spend stops. No alert-and-hope.
Chargeback and showback
Exportable cost data by team, project, or customer. Drop it straight into your own BI tools or finance systems.

Forecast and anomaly alerts
See where spend is headed before it happens, and get flagged the moment a loop or spike deviates from the baseline. Nothing waits for month-end.

Auto Router, on top
Once spend is visible and capped, Auto Router optimizes it further, routing each request to the cheapest model that meets your quality bar. 10 to 40% savings, zero prompt rewrites.


Scoped access per team
No team touches what they shouldn't. Marketing gets its own tools, engineering gets theirs.

Guardrails on every tool call
The same rules that govern model calls apply to tool calls. Block, redact, or flag anything that shouldn't get through.

Full attribution and audit trail
See which agent called which tool and what it cost. Every tool call lands in your audit log.
Sovereignty
Air-gapped means air-gapped.
EU data residency by default. No US parent company, no CLOUD Act exposure. And when you need fully air-gapped, we mean it.
EU sovereignty
Managed, hosted in the EU. SEAL-4, the highest level on the sovereignty scale.
Your cloud
Your account, your VPC, your keys.


0ms
0ms
Gateway overhead, p99

0 req/s
0 req/s
Sustained throughput

0.00%
0.00%
Uptime, last 30 days







