Infrastructure · Chapter 03 / Foundation

The foundation
every AI system
eventually needs.

Compute, data planes, orchestration, observability and governance — engineered as one substrate. Bring any model. Deploy in your VPC. Run in twenty-eight regions.

28
regions
10⁵
GPUs orchestrated
SOC 2 · HIPAA
compliance
schematic · A-01rev 2026.07
L6Enterprise Applications
L5AI Services
L4Model Orchestration
L3Data Platform
L2Compute & GPU
L1Cloud Infrastructure
substrate · unified healthy
Architecture

Six layers. One substrate.

Every layer is independently versioned and governed. Nothing is a black box; every layer emits the same telemetry contract.

L6
Enterprise Applications

Copilots, decision surfaces and workflow apps consuming AI as an internal capability.

annotation · L6
8k+surfaces deployed
L5
AI Services

Agents, reasoning graphs and typed tool-use — orchestrated across multi-model runtimes.

annotation · L5
500msmedian model response
L4
Model Orchestration

Cost, latency and quality routing across hosted, open-weight and fine-tuned models.

annotation · L4
12concurrent models · per request
L3
Data Platform

Vector, hybrid and structured stores with change-data capture and RAG contracts.

annotation · L3
3.4 PBgoverned retrieval corpus
L2
Compute & GPU

H100 and MI300 pools across regions, scheduled by workload class and residency.

annotation · L2
10⁵GPUs orchestrated
L1
Cloud Infrastructure

Private VPC deployment, anycast egress, dedicated links and sovereign residency.

annotation · L1
99.99%platform availability
substrate · observability · governance · security all systems nominal
Manifesto

The strongest AI systems are invisible. What users experience is simplicity.
What powers it is exceptional infrastructure.

Cross-section

A single request
through every layer.

reference · A-02 · request pathrev 2026.07
L6 ingressL4 inferenceL3 dataClient · VPCprivate linkAI Gatewayauth · policyModel Routercost · qualitycr-reason-170B MoEcr-vision-222BOSS / OWvLLMGPU poolH100 · MI300Vector storepgvectorStructuredpostgresObjects3 · gcsKafka · CDCstreamobservability · evaluation · governance · security
ingress
12ms
orchestration
44ms
inference
218ms
response
232ms
Command Center

One operational surface for the whole substrate.

No juggling dashboards. Every region, model, agent, deployment and policy — visible from a single calm interface.

clickripple · infra · global
HEALTHY
Global fleet
98,412.GPUs · scheduled
last 60s
Utilization · 24h72%
Deploys · 24h
142
+8
Failovers
3
0 alerts
Mean age
2h 41m
canary
Live activity
  • 00:14deploy
    cr-reason-1 · v14 · canary 34%
  • 00:11scale
    eu-central · +12 H100
  • 00:07failover
    ap-tokyo → ap-singapore · 48s
  • 00:02policy
    residency check · PASS
  • -01:32index
    vector rebuild · shard 08
Region health
27 nominal · 1 elevated
substrate · v2026.7.03p95 218ms · queue 4 · deploys 142
Global Fabric

28 regions. One control plane.

Pin workloads to a sovereign region. Route to the nearest healthy replica. Fail over in seconds without operator intervention.

us-westus-easteu-westeu-centralap-tokyoap-singapore
Governance

The controls that let an enterprise say yes.

Control
Zero-trust perimeter
mTLS between services, workload identity, per-request policy evaluation.
Control
Data residency
Pin every workload — training, inference, telemetry — to a specific region and boundary.
Control
OpenTelemetry native
Traces, metrics and evals emitted in one schema, shipped anywhere.
Control
Compliance ready
SOC 2 Type II, HIPAA, ISO 27001, GDPR, DORA — evidence generated automatically.
Service Level

Commitments in ink.

Multi-region active-active
99.99%
Control-plane uptime
Measured at edge, per region
p95 < 250ms
Inference latency
Automatic, without operator
< 60s
Failover
Named humans, not tiers
24×7
On-call engineering