Why Kafka?
Kafka as a distributed log: offsets, partitions, keys, replication and replay.
Transcript
Kafka is a distributed log: events are appended and kept, not removed when they are read.
Each consumer tracks its own offset, so many consumers can read the same events at their own pace.
A topic is split into partitions, so writes and reads spread across many machines.
Events with the same key always go to the same partition, so their order is kept.
Partitions are replicated across brokers, so the data survives a machine failure.
Because events are kept, you can replay history: rebuild a service, or feed a new consumer from the start.
Kafka is a great fit for event streaming, analytics pipelines, and event sourcing.
Append. Partition. Replay.
