r/apachekafka
How do you all catch a stuck consumer or growing lag before it blows up in prod?
- upvotes
- 1
- comments
- 10
Post
Highlighted: the lines this signal was extracted from
Been thinking about this a lot lately ... on Confluent Cloud specifically, what do you actually use to know when a consumer group is falling behind or straight up stuck? Feels like right now it's either roll your own Grafana/Prometheus/Burrow setup, or pay a ton for Datadog/Conduktor. Anyone found a middle ground that doesn't eat half a day to set up? What do you use today and what annoys you about it?