configpolicy

Author	SHA1	Message	Date
Dustin C. Hatch	2a110d7aba	hosts: Deploy haproxy0 _haproxy0.pyrocufflink.blue_ is a Fedora Linux VM that runs HAProxy to provide reverse proxy, exposing web sites and applications to the Internet. It has a static MAC address because it will need a static IP address, at least initially, in order for DNAT to work.	2024-08-24 11:46:40 -05:00
Dustin C. Hatch	2fa28dfa5f	r/dch-proxy: Update and clean up The dch-proxy role has not been used for quite some time. The web server has been handling the reerse proxy functionality, in addition to hosting websites. The drawback to using Apache as the reverse proxy, though, is that it operates in TLS-terminating mode, so it needs to have the correct certificate for every site and application it proxies for. This is becoming cumbersome, especially now that there are several sites that do not use the _pyrocufflink.net_ wildcard certificate. Notably, Tabitha's _hatchlearningcenter.org_ is problematic because although the main site are hosted by the web server, the Invoice Ninja client portal is hosted in Kubernetes. Switching back to HAProxy to provide the reverse proxy functionality will eliminate the need to have the server certificate both on the backend and on the reverse proxy, as it can operate in TLS-passthrough mode. The main reason I stopped using HAProxy in the first place was because when using TLS-passthrough mode, the original source IP address is lost. Fortunately, HAProxy and Apache can both be configured to use the PROXY protocol, which provides a mechanism for communicating the original IP address while still passing through the TLS connection unmodified. This is particularly important for Nextcloud because of its built-in intrusion prevention; without knowing the actual source IP address, it blocks _everyone_, since all connections appear to come from the reverse proxy's IP address. Combining TLS-passthrough mode with the PROXY protocol resolves both the certificate management issue and the source IP address issue. I've cleaned up the _dch-proxy_ role quite a bit in this commit. Notably, I consolidated all the backend and frontend definitions into a single file; it didn't really make sense to have them all separate, since they were managed by the same role and referred to each other. Of course, I had to update the backends to match the currently-deployed applications as well.	2024-08-24 11:46:28 -05:00
Dustin C. Hatch	cd1d472b74	scripts: Add VM host maintenance scripts The `migrate-all.sh` script is used to migrate one or more VMs (default: all) from one VM host to the other on demand. The `shutdown-vmhost.sh` script prepares a VM host to shut down by evicting Kubernetes Pods from the Nodes running on that host and then shutting them down, followed by migrating the rest of the running VMs to the other host.	2024-08-23 09:43:24 -05:00
Dustin C. Hatch	779b4e6a60	vault: Add missing DKMS variables	2024-08-23 09:39:54 -05:00
Dustin C. Hatch	d6a92a1151	websites: Fix file mode Playbook files should not be marked executable.	2024-08-23 09:34:25 -05:00
Dustin C. Hatch	aab581e859	hosts: Move VM hosts from hosts.offline Originally, the VM hosts were in a separate inventory so they would not be managed with the rest of the servers. It used to be that one server was running all the VMs, while the other was asleep. That's no longer the case; both alre always running and each has about half of the VMs. Since they're both always online, they can be managed normally now.	2024-08-23 09:33:29 -05:00
Dustin C. Hatch	153b210a73	vm-hosts: Do not reboot after auto updates For obvious reasons, the VM hosts cannot automatically reboot themselves.	2024-08-23 09:33:29 -05:00
Dustin C. Hatch	4aa06a02d6	hosts: Add vm-hosts to promtail group Enable collecting logs from VM hosts.	2024-08-23 09:33:29 -05:00
Dustin C. Hatch	5376253c9a	r/vmhost: Ensure vm-autostart waits for NFS mount If _vm-autostart.service_ starts before `/var/lib/libvirt/images` has been mounted, starting (some) virtual machines may fail.	2024-08-23 09:33:29 -05:00
Dustin C. Hatch	c546f09335	smtp-relay: Rewrite dustin@hatch.name Sometimes, the mail server for hatch.name is extremely slow. While there isn't much I can do about it for external senders, I can at least ensure that email messages sent by internal services like Authelia are always delivered quickly by rewriting the recipient address to my actualy email address, bypassing the hatch.name exchange entirely.	2024-08-22 16:17:00 -05:00
Dustin C. Hatch	0e46599fbc	r/postfix: Support rewriting recipient addresses The postfix role will now generate configuration and a lookup table for [canonical address mapping][0] of email recipients. To configure the mapping, the `postfix_recipient_canonical_map` must be a dictionary of source-target addresses, e.g.: ```yaml postfix_recipient_canonical_map: my.bad.email@fake.test: my.real.email@example.com ``` [0]: https://www.postfix.org/ADDRESS_REWRITING_README.html#canonical	2024-08-22 16:17:00 -05:00
Dustin C. Hatch	d7a271de20	Merge branch 'feature/redeploy-frigate'	2024-08-14 20:30:48 -05:00
Dustin C. Hatch	6e5e12f8b6	hosts: Add nvr2.p.b to collectd-sensors group To enable collecting temperature et al. sensor data.	2024-08-14 20:26:11 -05:00
Dustin C. Hatch	a2cf78f3f5	vm-hosts: Update vm-autostart logs0.pyrocufflink.blue has been replaced by loki0.pyrocufflink.blue since ages, so I'm not sure how I hadn't updated the autostart list with it yet. unifi3.pyrocufflink.blue replaced unifi2.p.b recently, when I was testing Luci/etcd.	2024-08-14 20:26:11 -05:00
Dustin C. Hatch	6d65e0594f	frigate: Configure HTTPS proxy with creds Only the _frigate_ user is allowed to access the Github API via the proxy.	2024-08-14 20:26:11 -05:00
Dustin C. Hatch	14a7d39e11	gw1/squid: Allow Frigate access to Github API Frigate uses the Github API to check for new releases. It then populates the `update.frigate_server` entity in Home Assistant via MQTT with the information it retrieved. If it is unable to access the Github API, the Home Assistant entity will be marked as "unavailable," which triggers an alert notification from Home Assistant. Thus, we need to allow Frigate to access Github if we want to use that entity as an indicator of whether or not Frigate is connected to the MQTT broker. I don't want to allow access to the Github API to everything on the Frigate server, just Frigate itself. To do that, I've assigned a unique username and password for Frigate. Only requests with the proper `Proxy-Authorization` header will be allowed access. By providing the credentials only the Frigate container, we can ensure no other process has access. I think I did this mostly as an exercise; there's no particular reason to disallow access to the Github API, since it's mostly read-only and can't really be used to exfiltrate any data (probably?).	2024-08-14 20:26:11 -05:00
Dustin C. Hatch	0ec9401c6e	r/squid: Support configuring auth_param The `auth_param` directive is used to configure proxy authentication helpers.	2024-08-14 20:26:11 -05:00
Dustin C. Hatch	27b172f083	r/system-auth: skip session winbind for local users If winbind is unable to communicate with any domain controller, the `pam_winbind.so` module will time out. In _auth_ and _account_ context, this was not an issue, at least for local users, because other modules terminated the stack before `pam_winbind.so` was called. In _session_ context, though, nothing terminated the stack at all, so `pam_winbind.so` was called unconditionally. This prevented even _root_ from logging in on the console. This made troubleshooting difficult, especially for the VM hosts, when the domain controllers were down.	2024-08-13 21:04:42 -05:00
Dustin C. Hatch	9ae88a5f36	datavol: Only set SELinux label of root directory Restoring the SELinux label of a mount point is really only necessary for a band new filesystem, which will have no label at all. In other cases, changing the context is probably neither necessary nor desirable, as the existing data is potentially labelled correctly already. Changing the label on on only the root directory should be sufficient to ensure applications run correctly with newly-provisioned filesystems, since they only have one directory anyway, without impacting too much for existing filesystems.	2024-08-12 22:22:50 -05:00
Dustin C. Hatch	34fcaa52ef	r/frigate: Increase service startup timeout Starting the Frigate container can take quite some time, since `podman` needs to check permissions/SELinux labels for the entire media volume.	2024-08-12 22:22:50 -05:00
Dustin C. Hatch	217b13faff	dch-root-ca: Add PB to trust DCH Root CA For machines that are not members of the AD domain	2024-08-12 22:22:50 -05:00
Dustin C. Hatch	d2b3b1f7b3	hosts: Deploy production Frigate on nvr2.p.b nvr2.pyrocufflink.blue originally ran Fedora CoreOS. Since I'm tired of the tedium and difficulty involved in making configuration changes to FCOS machines, I am migrating it to Fedora Linux, managed by Ansible.	2024-08-12 22:22:50 -05:00
Dustin C. Hatch	6c71d96f81	r/frigate-caddy: Deploy Caddy in front of Frigate Deploying Caddy as a reverse proxy for Frigate enables HTTPS with a certificate issued by the internal CA (via ACME) and authentication via Authelia. Separating the installation and base configuratieon of Caddy into its own role will allow us to reuse that part for other sapplications that use Caddy for similar reasons.	2024-08-12 18:47:04 -05:00
Dustin C. Hatch	59be10a51c	r/gasket-dkms: Build/sign Coral TPU driver The gasket-dkms package provides the `gasket` and `apex` kernel modules, which are needed fro the Google Coral Edge TPU. Since these are out-of-tree modules, they are not allowed in Fedora proper, so they are provided in a COPR, and have to be rebuilt for every kernel version. The DKMS framework handles automatically building the modules whenever the kernel updates. For systems usign UEFI with SecureBoot enabled, kernel modules must be signed by a key trusted by the platform. For locally-built modules, we can use the Machine Owner Key (MOK). Unfortunately, enrolling a new MOK requires rebooting and manual intervention during the boot process. Therefore, the gasket-dkms role has a `pause` step to ensure someone is paying attention and able handle the key enrollment interactively. Eventually, I'd like to have an RPM package with these modules pre-built, so production servers do not need the kernel development tools (`perl`, `gcc`, headers, etc.). It will be tricky, though, to make sure the modules get rebuilt for every kernel version as Fedora releases them.	2024-08-12 18:47:04 -05:00
Dustin C. Hatch	3250628cd1	gw1/squid: Allow NVR servers access to repos The Frigate NVR servers, prod & test, need to be able to access Fedora COPR (for the gasket-dkms package) and Github Container Registry (for Frigate itself).	2024-08-12 18:47:04 -05:00
Dustin C. Hatch	8239b60634	newvm: Add --network argument Although the `newvm.sh` script had a variable to configure the value specified for the `--network` argument to `virt-install`, it didn't expose a way to set it. We need this ability so we can e.g. create VMs on non-default networks like `camera` or `mgmt`.	2024-08-12 18:47:04 -05:00
Dustin C. Hatch	8dfb2e3e4f	r/frigate: Clean up Frigate role * Switch to Quadlet-style `.container` for systemd unit * Update to new image tag naming scheme (not arch-specific) * Use environment variables for secrets * Allow the entire `frigate_config` variable to be overridden	2024-08-12 18:47:04 -05:00
Dustin C. Hatch	7b61a7da7e	r/useproxy: Configure system-wide proxy The useproxy role configures the `http_proxy` et al. environmet variables for systemd services and interactive shells. Additionally, it configures Yum repositories to use a single mirror via the `baseurl` setting, rather than a list of mirrors via `metalink`, since the proxy a) the proxy only allows access to _dl.fedoraproject.org_ and b) the proxy caches RPM files, but this is only effective if all clients use the same mirror all the time. The `useproxy.yml` playbook applies this role to servers in the needproxy group.	2024-08-12 18:47:04 -05:00
Dustin C. Hatch	f51e0fe2a9	r/samba-dc: Enable auto-restart for samba.service Setting `RestartSec` is not enough to enable auto-restart; the `Restart` option is also required, as it defaults to `no`.	2024-08-09 08:11:39 -05:00
Dustin C. Hatch	3214d4b9b2	gw1/squid: Allow UniFi controller to OCI registries The UniFi Network server needs to be able access the _linuxserver.io_/GitHub and Docker Hub OCI image registries for the Unifi Network and Caddy container images, respectively.	2024-07-31 18:41:13 -05:00
Dustin C. Hatch	805a900f8a	gw1/squid: Allow Invoice Ninja to Stripe API HLC uses Invoice Ninja Stripe integration to process credit card payments from parents.	2024-07-14 15:45:36 -05:00
Dustin C. Hatch	ed22f6311c	r/samba-dc: Auto restart samba Although it's rare, sometimes Samba crashes or fails to start. When this happens, restarting it is almost always enough to get it working again. Since all sorts of authentication problems can occur if one of the domain controllers is down, it's probably best to just have systemd automatically restart _samba.service_ if it ever stops for any reason.	2024-07-03 10:30:20 -05:00
Dustin C. Hatch	96bc8c2c09	vm-hosts: Update autostart list k8s-amd64-n0, k8s-amd64-n1, and k8s-amd64-n2 have been replaced by k8s-amd64-n4, k8s-amd64-n5, k8s-amd64-n6, respectively. db0 is the new database server, which needs to be up before anything in Kubernetes starts, since a lot of applications running there depend on it.	2024-07-03 08:52:15 -05:00
Dustin C. Hatch	afb7030e44	migration: Add PostgreSQL server migration script This script captures the steps taken to migrate from the PostgreSQL server in the Kubernetes cluster, managed by _postgres operator_, to the dedicated server on _db0.pyrocufflink.blue_. The data were restored from the backups created by _wal-e_, and then the new server was promoted to primary. Finally, I cleaned up the roles and databases that are no longer needed.	2024-07-02 20:45:12 -05:00
Dustin C. Hatch	4f202c55e4	r/postgres-exporter: Deploy postgres-exporter The [postgres-exporter][0] exposes PostgreSQL server statistics to Prometheus. It connects to a specified PostgreSQL server (in this case, a server on the local machine via UNIX socket) and collects data from the `pg_stat_activity`, et al. views. It needs the `pg_monitor` role in order to be allowed to read the relevant metrics. Since we're setting up the exporter to connect via UNIX socket, it needs a dedicated OS user to match the PostgreSQL user in order to authenticate via the _peer_ method. [0]: https://github.com/prometheus-community/postgres_exporter/	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	3f5550ee6c	postgresql: wal-g: Set PGHOST By default, WAL-G tries to connect to the PostgreSQL server via TCP socket on the loopback interface. Our HBA configuration requires certificate authentication for TCP sockets, so we need to configure WAL-G to use the UNIX socket.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	6caf28259e	hosts: db0: Promote to primary All data have been migrated from the PostgreSQL server in Kubernetes and the three applications that used it (Firefly-III, Authelia, and Home Assistant) have been updated to point to the new server. To avoid comingling the backups from the old server with those from the new server, we're reconfiguring WAL-G to push and pull from a new S3 prefix.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	090ebb0c1b	r/wal-g-pg: Schedule daily backups WAL archives are not much good without a base backup onto which they can be applied. Thus, we need to schedule WAL-G to create and upload a backup periodically.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	b83c6de28a	gw1/squid: Add more URLs for Fedora/CoreOS updates After adding these, unifi2.pyrocufflink.blue (FCOS) was finally able to update successfully.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	dfc1a36ee5	deploy.sh: Wrapper for deployment scripts The `deploy.sh` script ensures the execution environment is correct by configuring the Ansible Vault secret, unlocking the `rbw` vault, and requesting an SSH client certificate. It then runs the specified end-to-end deployment script from the `deploy` directory.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	2ce211b5ea	hosts: Add db0.p.b db0.pyrocufflink.blue will be the primary server in the new PostgreSQL database cluster. We're starting with Fedora 39 so we can have PostgreSQL 15, to match the version managed by the Postgres Operator in the Kubernetes cluster right now.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	d8472c64a2	wait-for-host: PB to wait for a host to come up This playbook just waits for a machine to become available. It's useful when running Ansible immediately after creating a new machine.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	5958904fa3	bootstrap: PB to bootstrap a new machine I've actually had this playbook for a _long_ time, just never bothered to commit it. It's useful for the very first time Ansible is run for a managed node to configure all the basic stuff.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	056548e3e0	newvm: Add script to run virt-install For the longest time, whenever I needed to create a new virtual machine, I just used `Ctrl+R` to find the last `virt-install` command I had run and tweaked it for the new machine. At some point, my `~/.zsh_history` overflowed, though, so the command I had been running got lost. To avoid this silliness in the future, I've created a script that runs `virt-manager` for me. As a bonus, it has some configuration flags for the parameters that often vary between machines. For most machines, though, the script can be run as simply as `newvm.sh name`.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	208fadd2ba	postgresql: Configure for dedicated DB servers I am going to use the postgresql group for the dedicated database servers. The configuration for those machines will be quite a bit different than for the one existing machine that is a member of that group already: the Nextcloud server. Rather than undefine/override all the group-level settings at the host level, I have removed the Nextcloud server from the postgresql group, and updated the `nextcloud.yml` playbook to apply the postgresql-server role itself. Eventually, I want to move the Nextcloud database to the central database servers. At that point, I will remove the postgresql-server role from the `nextcloud.yml` playbook.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	54ad68b5bb	datavol: Playbook to provision a data volume The `datavol.yml` playbook can provision one or more data volumes on a managed node, using the definitions in the `data_volumes` variable. This variable must contain a list of dictionaries with the following keys: * `dev`: The block device where the data volume is stored (e.g. `/dev/vdb`) * `fstype`: The type of filesystem to create on the block device * `mountpoint`: The location in the filesystem hierarchy where the volume is mounted * `opts`: (Optional) options to pass to the `mkfs` program when formatting the device * `mountopts`: (Optional) additional options to pass to the `mount` program when mounting the filesystem	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	edffaf258b	r/wal-g-pg: Deploy WAL-G for PostgreSQL This role installs `wal-g` from the DCH Yum repository, and creates a configuration file for it in `/etc/postgresql`. Additionally, it installs a custom SELinux policy module that allows `wal-g` to run in the `postgresql_t` domain (i.e. when spawned by the PostgreSQL server).	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	99c309240c	r/postgresql-cert: ACME certificates using certbot This role can be used to get a server certificate for PostgreSQL from an ACME CA using `certbot`. It fetches the initial certificate and copies it to the PostgreSQL configuration directory. It also sets up a post-renewal hook script that copies updated certificates and reload the server.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	9e742dc217	roles/postgresql-server: Rewrite role This rewrite brings a lot of improvements and new functionality to the postgresql-server role. The most noticeable change is the introduction of the `postgresql_config_dir` variable, which can be used to specify a different location for the PostgreSQL server configuration files, separate from the data storage directory. By default, the variable is set to `/etc/postgresql`. For some situations, it may be necessary to disable this functionality, which can be accomplished by setting the value of `postgresql_config_dir` to the same path as `pgdata_dir`. Note also that the `postgresql-setup` tool, and the corresponding `postgresql-check-db-dir` script, which are included in the Fedora/Red Hat distribution of PostgreSQL, do not support having separate configuration and data directories, so their use has to be disabled. Another significant improvement is to how the `postgresql.conf` file is generated. Any setting can be set now, using the `postgresql_config` variable; any key in this dictionary will be written to the configuration file. Note that configuration file syntax requires single quotes around string values, so these have to be included in the YAML value. To support deploying standby servers, the role now supports running a command to restore from a backup instead of running `initdb`. Additionally, the `postgresql_standby` variable can be set to `true` to create the `recovery.signal` file, configuring the server as a standby.	2024-07-02 20:44:29 -05:00
Dustin C. Hatch	93eeaaaed4	gw1: Allow access to DCH yum repo via proxy Allows installing _sshca-cli-systemd_ from Kickstart.	2024-06-26 18:39:25 -05:00

1 2 3 4 5 ...

919 Commits