The Runtime Theory
Cloud & InfrastructureIn production

Multi-Region Architecture: Active-Active, Failover, and Replication

Recording in progress
#multi-region#failover#replication#architecture

Multi-region is a list of trade-offs, not a diagram. We compare active-passive and active-active designs on the mechanics that matter: where writes land, how replication propagates, what the failover actually flips, and what RTO and RPO numbers really cost. We work through DNS-based failover, routing policies, cross-region replication of databases and object storage, and the conflict problem nobody draws: concurrent writes to two regions. The takeaway is grounded in latency math — what the speed of light does to synchronous replication and which designs cannot exist because of it.

Topics covered:

  • Active-passive vs. active-active failure domains
  • Cross-region replication paths for data stores
  • Failover: DNS, routing, and read-only fallback
  • Conflict handling and the latency floor of sync replication

Related articles

More in Cloud & Infrastructure

In production
cloud

Edge Computing Explained: Where Compute Actually Sits

What edge computing actually is — the compute tiers from device to far edge to cloud, latency and bandwidth budgets, and which workloads genuinely benefit.

Details
In production
cloud

Container Orchestration Basics: API Server, Controllers, and Scheduler

Container orchestration from first principles — what the API server, controller manager, and scheduler actually do, with Kubernetes as the working example.

Details
In production
cloud

The Cost of Distributed Systems: Coordination, Consistency, and Failure

What distributed systems actually cost — coordination, consistency, and failure taxes — quantified with quorum math, tail latency, and retry-storm dynamics.

Details
In production
cloud

DNS and Traffic Routing in the Cloud: From Resolver to Anycast

How DNS actually routes traffic in the cloud — record resolution, CDN anycast, geo routing, load balancer handoffs, and how TTLs shape your failover story.

Details
In production
cloud

Object Storage Under the Hood: PUT, GET, and Erasure Coding

What happens inside an object store — the PUT and GET paths, metadata partitions, erasure coding, and why object storage is eventually consistent.

Details
In production
cloud

Autoscaling Explained: The Controller Loop Behind Horizontal Scaling

How autoscaling actually works — the metrics window, desired-replica calculation, stabilization, and why naive CPU-based scaling oscillates under real load.

Details
In production
cloud

Serverless Cold Starts, Measured: Where the Latency Actually Goes

Measured cold-start latency across Lambda, Cloud Functions, and container runtimes — what actually takes time and which optimizations genuinely reduce it.

Details
In production
cloud

Kubernetes Scheduling Visualized: The Filter-Score Pipeline

What the Kubernetes scheduler actually does — how pending Pods become assigned to Nodes, the filter-then-score pipeline, and how taints and affinity shape placement.

Details
In production
cloud

Container Images Explained: Layers, Manifests, and Digests

How container images are actually built and run — layers, manifests, content digests, and what the runtime does at pull and run time.

Details

Depth, delivered weekly

One technical dispatch a week — articles and episode notes before they go public.

One technical dispatch per week. No noise.