PostgreSQL Replication Slot Retained WAL Alert


Use the Replication Slot Retained WAL alert in Mini DBA to monitor PostgreSQL instances and make this condition visible before it becomes a wider database incident.

Screenshot pending: Mini DBA PostgreSQL Replication Slot Retained WAL alert screenshot placeholder

Alert Summary

  • Platform: PostgreSQL
  • Alert category: Disk
  • Default enabled: true
  • Default evaluation frequency: Minute
  • Threshold label: GB retained
  • Unit: GB

What Mini DBA Checks

Mini DBA describes this alert as: Replication slots retaining too much WAL can fill the WAL volume when consumers lag or disappear. The default evaluation frequency is Minute, so the alert is intended to be close enough to operational reality for live triage.

Why This Alert Is Helpful

This alert protects storage and recovery capacity. It helps you find growth, retention, and backup conditions that can stop writes, break recovery objectives, or leave the server without enough working space for normal database activity.

When To Enable It

Enable it where replication, read replicas, failover, or reporting copies are part of the service design. Disable it on standalone systems that do not use replication so the alert list stays focused.

Threshold Guidance

Threshold meaning: GB retained. Major threshold: 50 GB. Minor threshold: 10 GB. Comparison direction: "over". Use higher thresholds on batch-heavy, development, or intentionally bursty systems where brief pressure is expected. Use lower thresholds on latency-sensitive production systems, small instances with little headroom, and services with strict recovery or availability commitments.

Remediation For An Active Alert

Free space by removing safe-to-delete files, expanding the volume or tablespace, moving growth-heavy objects, correcting retention settings, or shrinking only after a documented one-off event. For recovery-related areas, verify that backups and log shipping or archiving are healthy before deleting anything.

Investigation Workflow

  1. Confirm the alert is still active and note the first seen time, affected instance, and severity.
  2. Review the affected disk, tablespace, log stream, backup job, retention setting, and recent growth pattern in Mini DBA before changing configuration or ending sessions.
  3. Compare the current value with the normal baseline for the same time of day or maintenance window.
  4. Record the cause, corrective action, and whether thresholds or routing should be adjusted after the incident.

Avoiding Alert Noise

If this alert is noisy, check whether the workload naturally has short bursts. Raise the duration or threshold for expected batch windows, but keep a lower route for production systems where user-facing latency matters. Avoid simply disabling the alert until you know whether the noise is caused by threshold choice or by a real capacity trend.

Related Pages