VViktorinwarrenops.hashnode.dev·11h ago · 7 min readAlerting when a RabbitMQ queue has no consumersOriginally published at warrenops.io. A queue with messages and zero consumers is the quietest outage there is. Nothing errors. Producers publish happily. The dashboard shows the broker green. Then th00
YYuvraj.Rinbuild-break-fix-diaries-with-yuvrajj.hashnode.dev·6d ago · 5 min readPrometheus vs Grafana: What Does Each Tool Actually Do ? When I started learning DevOps and Kubernetes, I came across two tools very often: Prometheus and Grafana. At first, I was confused about what each tool actually does and why they are used together. 00
RBranga bashyaminranga-blog.hashnode.dev·Sep 19 · 16 min readNobody Cares About Your Stack Until Something BreaksSix things I learned writing and rewriting alert rules for systems that were already live, already watched, and still surprising people anytime On Grafana alerting, PromQL, and the difference between 10
AKAvyay Kachhiainavyay-dev.hashnode.dev·Sep 18 · 18 min readBuilding PrismQTC: How I Built an Autonomous B2B Quote-to-Cash & Pricing Governance Platform Most B2B software treats sales as a sequence of isolated, static records: a quotation generated as an un-editable PDF, followed by disconnected email chains, manual margin approvals in spreadsheets, d00
MTMRIDUL TIWARIinmriduliti.hashnode.dev·Sep 12 · 7 min readWe tagged EKS workers like app servers — and Prometheus started scraping telegraf on nodes that never had itThe Telegraf Down alerts started piling up on a Thursday morning, and at first glance it looked bad. Nine targets, all critical, all under the same rule name. My first instinct was a fleet-wide agent 00
SKSushanth Kamabathulainblundersnbuilds.hashnode.dev·Sep 11 · 3 min readPostHog vs. Prometheus and Grafana Was a Category MistakeThey weren't competing for the same job, and I spent longer than I should have comparing them like they were. I spent a while trying to figure out which telemetry stack to pick for a side project, Pos00
PSPuneet Singhinfreecodecamp.org·Sep 4 · 29 min readClaude Code Observability with OpenTelemetryAgentic coding tools like Claude Code, OpenAI Codex, Google Antigravity, and Cursor have become ubiquitous for everyday software development. As agentic systems mature, much of the work developers hav10
GYGulshan Yadavinmrgulshanyadav.hashnode.dev·Aug 28 · 5 min readOpenTelemetry vs Prometheus: Why You Almost Certainly Need BothThis gets framed as a choice, and it mostly isn't one. They occupy different layers, and a normal production setup runs both. Understanding why saves a lot of wasted migration effort. What each one ac00
LSLiviu Stirbinblog.blockingqueue.com·Aug 27 · 10 min readA short crash course on metrics: gauges, counters, and histogramsThe metric type you pick is a promise about how the number behaves over time. Pick the wrong one and everything you build on top of it breaks silently. No error, no warning, just a graph that stops ma00
PUPurity Udehinfreecodecamp.org·Aug 22 · 10 min readHow to Convert Prometheus Histograms to OTLP with the OpenTelemetry CollectorModern applications often expose metrics at a /metrics endpoint using the Prometheus format. Among these metrics, histograms are particularly useful. They show how often values fall into different ran00