← All Series

SERIES // 11 PHASES · 60 ARTICLES · IN PROGRESS

Consensus Algorithms —
From First Principles to QuePaxa & Meerkat.

A book-quality series for engineers, not a paper list. Starts at "why does consensus exist" — replication, time, and failure — and walks the entire lineage: 2PC, Paxos, Raft, Byzantine fault tolerance, the modern Paxos family (EPaxos, Flexible Paxos, Atlas, Caesar, OmniPaxos), how etcd/ZooKeeper/Spanner/CockroachDB actually use it, and the current research frontier — QuePaxa and Cloudflare's Meerkat. By the last article you can read an SOSP/OSDI/NSDI consensus paper without translation.

60 articles live
60 planned
11 phases
11 / 11 phases done


Phase 1

Foundations

Why consensus exists in the first place — before any algorithm. Distributed systems, the replication problem, CAP done properly, and what comes after CAP.

Phase 2

Time and Failure

There is no "now" in a distributed system. Physical and logical clocks, failure detection, and the impossibility result that explains why every consensus algorithm makes a specific compromise.

Phase 3

Replication

The machinery consensus sits on top of — state machine replication, logs, quorums, and how a "read" can be made safe without asking everyone every time.

Phase 4

Consensus Basics

Safety vs. liveness, leader election, atomic broadcast — and the two commit protocols that predate real consensus and explain exactly why they aren't enough.

Phase 5

Classical Consensus — Paxos & Raft

The algorithms everything else is measured against: single-decree and Multi-Paxos, Viewstamped Replication, Raft end to end, and Zab.

Phase 6

Byzantine Consensus

When nodes don't just crash — they lie. The Byzantine Generals Problem, PBFT, and the modern BFT lineage that powers blockchains.

Phase 7

The Modern Paxos Family

Where this series stops being "another Raft explainer." Leaderless consensus, flexible quorums, and the research lineage that leads directly to QuePaxa and Meerkat.

Phase 8

Consensus in Production

How the algorithms show up in systems you already run — etcd, ZooKeeper, Consul, Kafka's KRaft, Kubernetes — and how Jepsen actually tests the claims.

Phase 9

Cloud-Native & Geo-Distributed Consensus

Consensus stretched across continents. Spanner's TrueTime, CockroachDB's Multi-Raft, FoundationDB, TiDB/YugabyteDB, and the WAN-latency tricks that make it bearable.

Phase 10

Research Papers

The frontier, read paper-by-paper in a consistent format: problem, core idea, architecture, failure handling, performance, trade-offs. Ending at QuePaxa and Meerkat.

Phase 11

Future Directions & Capstone

Where consensus is headed, and the payoff: build your own consensus algorithm, benchmark it against Raft, and walk away with a working mental map of the entire field.