The Stack
- Grafana for dashboards and visualization
- Prometheus for metrics storage and PromQL queries
- Loki for log aggregation and LogQL queries
- Tempo for distributed traces
- Alloy as the collector that receives, processes, and forwards signals
Bug 1: The Loki Pipeline That Looked Healthy
Bug 2: Dashboards That Reference Metrics That Do Not Exist
Bug 3: The Phantom Cost Panel
Bug 4: Counter Used Where a Gauge Was Needed
Bug 5: Datasource UIDs vs Placeholder Names
Bug 6: Warning Panels That Matched Words Instead of Severity
What I Actually Measure Now
Tech Stack
- Grafana for dashboards and visualization
- Prometheus for metrics storage and PromQL
- Loki for log aggregation and LogQL
- Tempo for distributed traces
- Alloy for metric, log, trace, and Kubernetes telemetry pipelines
- node_exporter for host metrics
- kube-state-metrics for Kubernetes object state
- Docker Compose for the central LGTM stack
- Kubernetes for in-cluster telemetry collection and service discovery