AI Gateway HQ vs. TrueFoundry AI Gateway
A focused AI gateway with conservative cash controls and an administration model a sysadmin can operate without adopting a broader ML platform.
- Pre-dispatch control
- Operator-ready evidence
- Prompt storage off by default
Choose AI Gateway HQ when a focused, quickly administered AI control boundary is preferable to a larger platform purchase. The product's differentiators are strict financial exposure control, minimal-data evidence, and guided operations; TrueFoundry remains relevant when the broader platform is itself the goal.
Routes, keys, policies, budgets, request evidence, support grants, and billing are explicit objects in one console; strict cost reservations occur before forwarding rather than relying only on reported consumption.
Compare the operating boundary.
AI Gateway HQ entries describe implemented product behavior. Alternative entries summarize the linked first-party documentation—not anonymous review scores.
One OpenAI/Anthropic-compatible endpoint; encrypted multi-account BYOK pools; capability-first priority, weighted, request-cost, health, and request-aware provider-capacity selection; shared quota cooldowns and bounded, reason-coded fallback.
Universal API with weighted, latency, and priority routing; fallback, caching, MCP, observability, and broad provider support.
Atomic organization-and-workload reservation before forwarding, strict rate and concurrency enforcement, explicit output caps, and settlement against supported provider-reported usage. Promotional credit cannot fund server-paid model exposure.
Budgets, rate limits, virtual keys, and model/provider controls across platform scopes.
OIDC administration, mandatory MFA, built-in least-privilege roles, virtual workload keys, signed execution context, Observe/Shadow/Enforce policy, and local jailbreak, injection, exfiltration, encoding, and Unicode risk signals.
Extensive input/output guardrail hooks, RBAC, SSO, and audit features; broader enterprise governance today.
WAF-protected AWS serverless deployment, tenant-bound KMS encryption, signed releases, payload-free request metadata by default, tamper-evident audit exports, and customer-approved time-bounded support access.
Managed, VPC, on-premises, air-gapped, and multi-region enterprise positioning.
A free BYOK proving tier, then $0.10 per 1,000 successful Flex requests with no percentage markup on inference purchased through customer-owned provider accounts; higher-control plans are scoped by operating requirements.
Developer is free for 50,000 requests; Pro lists $499/month, Pro Plus $2,999/month, and Enterprise is custom, with published user/request allowances.
Consider TrueFoundry when its larger AI platform, private deployment menu, distributed or semantic caching, or current guardrail integration catalog is already required.
Facts were reviewed from the linked first-party documentation and pricing pages on August 7, 2026. Public meters are not normalized: requests, logs, credits, infrastructure, and enterprise capacity are different units. Revalidate pricing and capabilities before purchasing.