Operations and reliability teams use this concept when setting objectives, investigating incidents, improving resilience, and communicating service health.
Related concepts include Observability.
Additional Resource: CNCF Cloud Native Glossary.
Site Reliability Engineering and operations teams use this concept to measure service health, investigate incidents, and turn operational learning into improvement.
Related concepts include DevOps and Incident Response.
Additional Resource: Google Site Reliability Engineering Book.