Proxmox VE: Remove a Dead Node and Go Back to Standalone

1. Diagnose

pvecm nodes                                     # cluster members
pvecm status                                    # quorum
ls -l /etc/pve/nodes                            # node folders
ls -l /etc/pve/nodes/DEAD_NODE/qemu-server      # VMs defined on the dead node

2. Remove the dead node

Only if it is never coming back with its current configuration.

pvecm expected 1              # regain quorum with a single surviving node (2-node cluster)
pvecm delnode DEAD_NODE

If its VMs had disks on shared storage, move their configs to a live node before cleaning up:

mv /etc/pve/nodes/DEAD_NODE/qemu-server/*.conf /etc/pve/nodes/LIVE_NODE/qemu-server/
rm -rf /etc/pve/nodes/DEAD_NODE

3. Go back to standalone (optional)

To dissolve the cluster on the surviving node:

systemctl stop pve-cluster corosync
pmxcfs -l                     # local mode, lets you edit without quorum
rm /etc/pve/corosync.conf
rm -rf /etc/corosync/*
killall pmxcfs
systemctl start pve-cluster

A reboot at the end doesn't hurt. If you use Ceph, handle its monitors and OSDs separately.

Docs: Proxmox cluster manager.

Written by Daniel Ruiz Peláez, Systems & Infrastructure Engineer (Linux, VMware, Proxmox, Active Directory, networking and security). These are notes from real problems I have solved.

← Back to all posts