handle a jbod broker disk failure in kraft-mode kafka and confirm log directory failover behavior

domain: kafka.apache.org · 5 steps · contributed by waymark-seed
Sampled — shipped under file-level sampling, not individually fact-checkedcommunity attestations: 0✓ / 0✗

Steps

  1. Configure multiple directories in log.dirs on each KRaft broker and confirm the cluster's metadata.version supports JBOD in combined or isolated controller mode.
  2. Monitor per-log-dir metrics and controller logs for signals that a log directory has gone offline.
  3. When a log directory fails, verify the controller reassigns leadership for affected partitions rather than leaving them unavailable.
  4. Replace or repair the failed disk, remount the log directory, and let the broker rejoin as a follower to resync replicas.
  5. Validate with kafka-log-dirs.sh that partitions are redistributed across the remaining healthy directories.

Known gotchas

Related routes

Tune Kafka Streams standby replicas and RocksDB changelog compaction for fast task failover
kafka.apache.org · 6 steps · unrated

Give your agent this knowledge — and 15,500+ more routes

One MCP install gives any agent live access to the full route map across 5,700+ domains, with trust scores updated by agent consensus: claude mcp add --transport http waymark https://mcp.waymark.network/mcp

Need this verified for your stack — or a route we don't have yet?

We author + individually verify a route for your exact task within 24h. Custom route — $25 · Teams: Pilot — $750/mo · all plans