Observability

7 guides tagged “Observability”.

  1. How to Set Up Basic Application Monitoring Set up basic app monitoring with Prometheus metrics, a health endpoint, and Grafana dashboards. A tool-agnostic tutorial with a concrete example. DevOps beginner 8 min
  2. What Is Log Rotation (and Why Do Your Logs Keep Disappearing)? Log rotation archives and deletes old logs so disk never fills up. Learn how it works, why your logs vanish, and how log management services fit in. DevOps beginner 5 min
  3. What Is Observability (and How Is It Different From Monitoring)? Monitoring tells you when something known is wrong; observability lets you ask new questions about unknown failures. Learn the difference with examples. DevOps beginner 7 min
  4. What Are SLA, SLO, and SLI? SLIs are the measurements, SLOs are your internal targets, and SLAs are contractual promises to customers. Learn the difference with a comparison table. DevOps beginner 6 min
  5. What Is SRE (Site Reliability Engineering)? Site reliability engineering applies software engineering to operations: SLOs, error budgets, and automation. Learn the core ideas and key practices. DevOps beginner 7 min
  6. What Is Self-Healing Infrastructure? Self-healing infrastructure automatically detects failures and restores the desired state without human action. Learn the patterns and the risks. DevOps intermediate 6 min
  7. What Is AIOps? AIOps applies machine learning to IT operations to reduce alert noise, detect anomalies, and speed up incident response. Learn what it does and its limits. DevOps intermediate 6 min