Segments

DevOps & Platform Engineering

DevOps and Platform Engineering are approaches aimed at improving collaboration between software development and IT operations.

Model order
  1. Knowledge domains
  2. /Thematic areas
  3. /Segments
  4. /Building blocks
View
Segment
Type
Classification

Reliability & Observability

This segment covers concepts and approaches for ensuring stable and observable system behavior during operation. It includes mechanisms for measuring and evaluating reliability, capturing states and events, and analyzing deviations and failures. It describes how platforms and services are made observable in order to understand and interpret their behavior over time, without focusing on infrastructure definition or delivery processes. Topics related to platform provisioning, software delivery, or business logic are addressed in other segments.

MethodReliability & Observability

Root Cause Analysis (RCA)

A structured approach to identify the root causes of problems.

#Product#Delivery
ConceptReliability & Observability

Metrics

Metrics help measure and analyze the performance and efficiency of processes.

#Data#Analytics
ConceptReliability & Observability

Observability

Observability enables understanding the state of complex systems through metrics, logs, and traces.

#Observability#Reliability
ConceptReliability & Observability

Reliability

Reliability is a critical concept in system development that ensures systems consistently perform as expected.

#Observability#Reliability
ConceptReliability & Observability

Service Level Agreement (SLA)

A Service Level Agreement (SLA) defines the expectations for the services provided by a vendor.

#Observability#Reliability
ConceptReliability & Observability

Service Level Indicator (SLI)

A Service Level Indicator (SLI) measures the quality of a service against predefined criteria.

#Observability#Reliability
ConceptReliability & Observability

Service Level Objective (SLO)

A Service Level Objective (SLO) defines specific performance expectations for a service.

#Observability#Reliability
ToolReliability & Observability

Grafana

Grafana is an open-source tool for visualizing and analyzing data.

#Data#Platform
TechnologyReliability & Observability

ELK Stack (Elasticsearch, Logstash, Kibana)

The ELK Stack combines Elasticsearch, Logstash, and Kibana for efficient data processing and visualization.

#Data#Analytics
TechnologyReliability & Observability

Prometheus

Prometheus is an open-source monitoring and alerting system.

#Data#Analytics