← All runbooks in Observability
Runbook: Observability Storage Full
1 · Prerequisites
Confirm every item is in place before any state change.
- Prometheus / Loki / Tempo
- Storage
2 · Pre-checks
Read-only diagnostic commands. If any of these don't match expected output, stop and investigate further.
- · Identify which storage is full
- · Identify growth pattern
3 · Procedure
Execute each step in order. Verify the expected output of a step before moving to the next.
- 1Prometheus: check TSDB size
- 2Loki: check object store size
- 3Tempo: check object store size
- 4If Prometheus: increase retention size or scale
- 5If Loki / Tempo: prune old chunks / blocks or scale object storage
- 6If growth is anomalous: investigate the source
- 7Validate
4 · Verification
Confirm the procedure actually fixed the problem.
- ✓Storage has free capacity
- ✓Retention / ingestion rate is normal
- ✓No data loss expected
5 · Rollback
If verification fails, undo the procedure in reverse order.
- ↶Scale storage back to the previous configuration after root cause is fixed
6 · Escalation
When the runbook isn't enough, contact:
- · Coordinate with the storage owner
Purpose
Observability Storage Full
When to use this runbook
Use this runbook when the operator needs a guided procedure to handle the situation described above.
Pre-checks
Before starting the procedure, confirm the prerequisites and pre-checks are met. The structured lists are rendered from the frontmatter by the page layout.
Procedure
Follow the steps from the frontmatter procedure steps. The page layout renders the steps as a checklist with copy-to-clipboard affordances.
Verification
After the procedure, the structured verification items from the frontmatter are rendered as a checklist.
Rollback
If the procedure fails or makes things worse, follow the structured rollback steps from the frontmatter.
Escalation
The structured escalation path is rendered from the frontmatter. Use it if the operator cannot complete the procedure safely.