Observability & Analysis

Learning Notes #

This is my learning Notes for Monitor/Metrics/Tracing

Monitoring #

NameDocsCodesCommentsNext
OpenTelemetryDocsGithubOpenTelemetry is a collection of APIs, SDKs, and tools for generating, collecting, and exporting telemetry data (traces, metrics, and logs).Getting Started with OpenTelemetry
PrometheusDocsGithubPrometheus is a systems and service monitoring system.
ThanosDocsGithubThanos is a set of components that can be composed into a highly available Prometheus setup.
ZabbixDocsGithubZabbix is an enterprise-class open-source monitoring solution for tracking the status of various network services, servers, and other network hardware.
GrafanaDocsGithubGrafana is an open source, feature rich metrics dashboard and graph editor for Graphite, InfluxDB, Prometheus and other time series databases.
ZipkinDocsGithubZipkin is a distributed tracing system.
FluentdDocsGithubFluentd is an open source data collector for unified logging layer.
KibanaDocsGithubKibana is an open source data visualization plugin for Elasticsearch.
LokiDocsGithubLoki is a horizontally-scalable, highly-available, multi-tenant log aggregation system inspired by Prometheus.

Prometheus #

Architecture

Demo with Grafana


Learning #

Introduction to Prometheus

PromLabs

Prometheus Course Outline

Chapter 1. Course Introduction

Chapter 2. Introduction to Observability

Chapter 3. Introduction to Prometheus

Chapter 4. Installing and Setting Up Prometheus

Chapter 5. Basic Querying

Chapter 6. Dashboarding

Chapter 7. Monitoring Host Metrics

Chapter 8. Monitoring Container Metrics

Chapter 9. Instrumenting Code

Chapter 10. Building Exporters

Chapter 11. Advanced Querying

Chapter 12. Relabeling

Chapter 13. Service Discovery

Chapter 14. Blackbox Monitoring

Chapter 15. Pushing Data

Chapter 16. Alerting

Chapter 17. Making Prometheus Highly Available

Chapter 18. Recording Rules

Chapter 19. Scaling Prometheus Deployments

Chapter 20. Local Storage

Chapter 21. Remote Storage Integrations

Chapter 22. Transitioning From and Integration with Other Monitoring Systems

Chapter 23. Monitoring and Debugging Prometheus

Chapter 24. Prometheus and Kubernetes


Tracing #


Chaos Engineering #


Reference #