Glossary

Observability

Observability is the property of a system that allows its internal state to be understood from the signals it emits, including for failures nobody anticipated.

Short answer

What is observability?

Observability is the property of a system that allows its internal state to be understood from the signals it emits, including for failures nobody anticipated.

Explained two ways

Plain and technical

In plain terms

Monitoring tells you that something is wrong, because someone decided in advance to watch for it. Observability is being able to work out why something is wrong even when nobody predicted that particular failure.

Technically

Observability is conventionally achieved through metrics, logs and traces, though having all three is not sufficient: they must share an identity model so that a metric leads to a trace and a trace leads to the logs emitted during it. The practical test is whether a new question about production behaviour can be answered without first deploying additional instrumentation.

Example

What it looks like in practice

Latency rises at the edge. An exemplar on the latency histogram links to a slow trace, the trace shows a database span taking four seconds, and that span links to the log lines showing which query it was. No new instrumentation was needed.
Keep going

More definitions

See these concepts in a running system

A 30-minute walkthrough against your own infrastructure rather than a slide about the theory.