Find a term
Understand service health, incidents, and the telemetry used to investigate production systems. Connect reliability outcomes to engineering decisions.
Terms beginning with O
9 termsObjective attainment
Objective attainment is a reliability concept used to describe a specific condition, control, or decision in the operation of software services.
Reliability and observabilityObservability
Observability is the ability to understand a system's internal behavior from the evidence it produces. In software systems, that evidence often includes logs, metrics, and traces connected to enough context to investigate unexpected behavior.
Reliability and observabilityObservability dashboard
Observability dashboard is a view combining metrics, logs, traces, and context for a system question.
Reliability and observabilityObservability signal
Observability signal is a recorded representation of system behavior that helps a team infer what is happening inside a service.
Reliability and observabilityOpenTelemetry
OpenTelemetry is an open-source framework and specification set for generating, collecting, and exporting telemetry.
Reliability and observabilityOpenTelemetry Collector
OpenTelemetry Collector is a vendor-neutral service that receives, processes, and exports telemetry.
Reliability and observabilityOperational toil
Operational toil is recurring work needed to keep a service running that tends to be manual, tactical, automatable, and without lasting improvement. Its volume often grows with the service unless the underlying need is reduced.
Reliability and observabilityOTLP
OTLP is the OpenTelemetry Protocol for transmitting telemetry between instrumented systems, collectors, and backends.
Reliability and observabilityOverload control
Overload control is a reliability concept used to describe a specific condition, control, or decision in the operation of software services.
Reliability and observability