- Capacity Planning for a Self-Hosted Observability Stack
A practical guide to sizing LGTM, Victoria, and ClickHouse for ingestion, retention, compression, compute, storage, network, and cost.
11 min read -
Designing a Telemetry Pipeline That Scales [Part 2]: Sampling, Kafka, Storage, and HAHow to make the pipeline durable at scale with tail sampling, Kafka buffering, VictoriaMetrics, Tempo, and production hardening.
14 min read -
Designing a Telemetry Pipeline That Scales [Part 1]: Architecture, Agents, and ForwardersWhy observability pipelines fail at 100 to 1,000 servers, and how to redesign collection and processing with OpenTelemetry.
26 min read