Learn Labs
10. Cross-Cluster Data Mirroring

10. Cross-Cluster Data Mirroring

Chapter 10 of Kafka: The Definitive Guide — 12 sections.

Source: Kafka: The Definitive Guide, 2nd Ed., Ch. 10 Terminology, established up front: "In most databases, continuously copying data between database servers is called replication. Since we've used replication to describe movement of data between Kafka nodes that are part of the same cluster, we'll call copying of data between Kafka clusters MIRRORING. Apache Kafka's built-in cross-cluster replicator is called MirrorMaker."

TermScopeMechanism
REPLICATIONwithin one cluster (Ch. 6/7)synchronous-ish, ISR-based
MIRRORINGbetween clusters (this chapter)ASYNCHRONOUS, consumer + producer

The easy case, dismissed immediately: "In some cases, the clusters are completely separated... different departments, different use cases, different SLAs/workloads, different security requirements. Those use cases are fairly easy — managing multiple distinct clusters is the same as running a single cluster multiple times." This chapter is about the interdependent case.