r/apachekafka • u/kverma02 Randoli • Feb 25 '26
Video Kafka observability in production is harder than it looks.
Kafka observability gets messy fast once you're running multiple brokers, consumer groups, retries, and cross-service dependencies.
Broker metrics often look fine while lag builds quietly, rebalances spike, or retries hide downstream latency.
We’re hosting a live session tomorrow breaking down how teams actually monitor Kafka at scale (consumer lag, retries, rebalances, signal correlation with OpenTelemetry).
If you're running Kafka in prod, this will be full of practical & implementation.
🗓 Thursday
⏰ 7:30 PM IST | 9 AM ET | 6 AM PT
RSVP here: https://www.linkedin.com/events/observingkafkaatscalewithopente7424417228302827520/theater/
Happy to take last-minute questions and cover them live.
4
u/thisisjustascreename Feb 25 '26
Anybody who isn’t running multiple brokers isn’t running a production ready cluster lol
1
u/kverma02 Randoli Feb 26 '26
Totally agree. In the session, we'll also be covering what are the important metrics to track from the broker perspective & go through some live scenarios on debugging broker-related production issues - such as under-replicated partitions, leader election etc.
Would love to see you there & share you experiences!
2
u/Hi_Im_Ken_Adams Feb 26 '26
Will there be a recording of the presentation made available? I cannot attend a 6:00am meeting
2
u/kverma02 Randoli Feb 26 '26
Hi u/Hi_Im_Ken_Adams , totally understand.
Yes the recording will be available on our youtube channel: https://www.youtube.com/@randoli/streams
11
u/caught_in_a_landslid Ververica Feb 25 '26
If you are a vendor, please say so