configpolicy

dustin

Author	SHA1	Message	Date
Dustin	164f3b5e0f	r/wal-g-pg: Handle versioned storage locations The target location for WAL archives and backups saved by WAL-G should be separated based on the major version of PostgreSQL with which they are compatible. This will make it easier to restore those backups, since they can only be restored into a cluster of the same version. Unfortunately, WAL-G does not natively handle this. In fact, it doesn't really have any way of knowing the version of the PostgreSQL server it is backing up, at least when it is uploading WAL archives. Thus, we have to include the version number in the target path (S3 prefix) manually. We can't rely on Ansible to do this, because there is no way to ensure Ansible runs at the appropriate point during the upgrade process. As such, we need to be able to modify the target location as part of the upgrade, without causing a conflict with Ansible the next time it runs. To that end, I've changed how the _wal-g-pg_ role creates the configuration file for WAL-G. Instead of rendering directly to `wal-g.yml`, the role renders a template, `wal-g.yml.in`. This template can include a `@PGVERSION@` specifier. The `wal-g-config` script will then use `sed` to replace that specifier with the version of PostgreSQL installed on the server, rendering the final `wal-g.yml`. This script is called both by Ansible in a handler after generating the template configuration, and also as a post-upgrade action by the `postgresql-upgrade` script. I originally wanted the `wal-g-config` script to use the version of PostgreSQL specified in the `PG_VERSION` file within the data directory. This would ensure that WAL-G always uploads/downloads files for the matching version. Unfortunately, this introduced a dependency conflict: the WAL-G configuration needs to be present before a backup can be restored, but the data directory is empty until after the backup has been restored. Thus, we have to use the installed server version, rather than the data directory version. This leaves a small window where WAL-G may be configured to point to the wrong target if the `postgresql-upgrade` script fails and thus does not trigger regenerating the configuration file. This could result in new WAL archives/backups being uploaded to the old target location. These files would be incompatible with the other files in that location, and could potentially overwrite existing files. This is rather unlikely, since the PostgreSQL server will not start if the _postgresql-upgrade.service_ failed. The only time it should be possible is if the upgrade fails in such a way that it leaves an empty but valid data directory, and then the machine is rebooted.	2024-11-17 10:27:31 -06:00
Dustin	8b9cf1985a	r/wal-g-pg: Schedule weekly delete jobs WAL-G slows down significantly when too many backups are kept. We need to periodically clean up old backups to maintain a reasonable level of performance, and also keep from wasting space with useless old backups.	2024-11-05 19:28:57 -06:00
Dustin	2e37fce4f6	r/wal-g-pg: Run wal-g backup as postgres `wal-g` needs to connect to the PostgreSQL database system, so it should run as the _postgres_ user, who has permission to connect, rather than _root_, who does not.	2024-08-30 09:44:43 -05:00
Dustin	090ebb0c1b	r/wal-g-pg: Schedule daily backups WAL archives are not much good without a base backup onto which they can be applied. Thus, we need to schedule WAL-G to create and upload a backup periodically.	2024-07-02 20:44:29 -05:00
Dustin	edffaf258b	r/wal-g-pg: Deploy WAL-G for PostgreSQL This role installs `wal-g` from the DCH Yum repository, and creates a configuration file for it in `/etc/postgresql`. Additionally, it installs a custom SELinux policy module that allows `wal-g` to run in the `postgresql_t` domain (i.e. when spawned by the PostgreSQL server).	2024-07-02 20:44:29 -05:00

5 Commits (a63ee2bff543a4cf7c3bc2386b758b920dca48f9)