Procedure

Add a node to a Proxmox cluster

Add a Proxmox VE node to an existing cluster by preparing DNS and hosts resolution, matching versions, validating Corosync networking and quorum, checking storage compatibility, joining the cluster, and validating pmxcfs and migrations.

Objective

Add a Proxmox VE node to an existing cluster by preparing DNS and hosts resolution, matching versions, validating Corosync networking and quorum, checking storage compatibility, joining the cluster, and validating pmxcfs and migrations.

Prerequisites

  • Known addressing, VLAN, routing and firewall context for the affected flow.
  • A representative test client and the exact protocol or port involved.
  • Administrative access appropriate to the system being changed or diagnosed.
  • A clearly identified scope: affected users, systems, addresses, services and the time of the observed problem.
  • A maintenance or test window when the procedure can affect production traffic or availability.
  • A copy of the current configuration or other recovery material before any irreversible action.

Step-by-step procedure

1

Establish the baseline and scope

Before changing anything, reproduce the issue or document the requested change on a representative system. Record the affected users or services, exact time, current configuration, recent changes and a known-good comparison point. This baseline is the reference used to decide whether each later step improves the situation.

Expected result
  • The scope and current state are documented well enough to reproduce or verify the procedure.
2

Preparing DNS

Implement this requirement in a controlled scope: preparing DNS. Capture the previous value or configuration first, make the smallest change that satisfies the design, and validate the effective state rather than assuming that a saved setting is active. If the expected result is not obtained, stop, restore the previous state, and reassess before expanding the change.

Expected result
  • Evidence for this area is explicit, reproducible and consistent with the intended design.
3

Hosts resolution

Implement this requirement in a controlled scope: hosts resolution. Capture the previous value or configuration first, make the smallest change that satisfies the design, and validate the effective state rather than assuming that a saved setting is active. If the expected result is not obtained, stop, restore the previous state, and reassess before expanding the change.

Expected result
  • Evidence for this area is explicit, reproducible and consistent with the intended design.
4

Matching versions

Implement this requirement in a controlled scope: matching versions. Capture the previous value or configuration first, make the smallest change that satisfies the design, and validate the effective state rather than assuming that a saved setting is active. If the expected result is not obtained, stop, restore the previous state, and reassess before expanding the change.

Expected result
  • Evidence for this area is explicit, reproducible and consistent with the intended design.
5

Corosync networking

Implement this requirement in a controlled scope: Corosync networking. Capture the previous value or configuration first, make the smallest change that satisfies the design, and validate the effective state rather than assuming that a saved setting is active. If the expected result is not obtained, stop, restore the previous state, and reassess before expanding the change.

Expected result
  • Evidence for this area is explicit, reproducible and consistent with the intended design.
6

Quorum

Implement this requirement in a controlled scope: quorum. Capture the previous value or configuration first, make the smallest change that satisfies the design, and validate the effective state rather than assuming that a saved setting is active. If the expected result is not obtained, stop, restore the previous state, and reassess before expanding the change.

Expected result
  • Evidence for this area is explicit, reproducible and consistent with the intended design.
7

Storage compatibility

Implement this requirement in a controlled scope: storage compatibility. Capture the previous value or configuration first, make the smallest change that satisfies the design, and validate the effective state rather than assuming that a saved setting is active. If the expected result is not obtained, stop, restore the previous state, and reassess before expanding the change.

Expected result
  • Evidence for this area is explicit, reproducible and consistent with the intended design.
8

Joining the cluster

Implement this requirement in a controlled scope: joining the cluster. Capture the previous value or configuration first, make the smallest change that satisfies the design, and validate the effective state rather than assuming that a saved setting is active. If the expected result is not obtained, stop, restore the previous state, and reassess before expanding the change.

Expected result
  • Evidence for this area is explicit, reproducible and consistent with the intended design.
9

Pmxcfs

Implement this requirement in a controlled scope: pmxcfs. Capture the previous value or configuration first, make the smallest change that satisfies the design, and validate the effective state rather than assuming that a saved setting is active. If the expected result is not obtained, stop, restore the previous state, and reassess before expanding the change.

Expected result
  • Evidence for this area is explicit, reproducible and consistent with the intended design.
10

Validate the complete service

Repeat the original user, system or application workflow from the real source and verify the complete result, not only one command or one local check. Confirm that logs and monitoring show the expected behavior and that no temporary debug, bypass, test account, rule or maintenance setting remains enabled.

Expected result
  • The end-to-end service works or the remaining failure is isolated to a clearly identified component.

Technical commands from the original procedure

These technical blocks are preserved byte-for-byte from the historical procedure and kept in their original order. Review names, addresses, paths and parameters before use.

Technical block 1
pvecm status
Technical block 2
pvecm nodes
Technical block 3
systemctl --no-pager --full status corosync pve-cluster
Technical block 4
hostnamectl
Technical block 5
cat /etc/hosts
Technical block 6
ip -br addr
Technical block 7
pveversion -v
Technical block 8
timedatectl
Technical block 9
ping -c 20 <CLUSTER-NODE-IP>
Technical block 10
ip route get <CLUSTER-NODE-IP>
Technical block 11
cp -a /etc/network/interfaces /root/interfaces.before-cluster
Technical block 12
pvesm status
Technical block 13
pvecm add <IP-NODE-EXISTANT>
Technical block 14
pvecm status
Technical block 15
pvecm nodes
Technical block 16
mount | grep /etc/pve
Technical block 17
ls -la /etc/pve/nodes
Technical block 18
pvesm status
Technical block 19
ip -br link
Technical block 20
qm migrate <VMID> <TARGET-NODE> --online
Technical block 21
pct migrate <CTID> <TARGET-NODE> --restart
Technical block 22
pvecm status

Validation

The procedure is validated when:

  • The original symptom or change request has been tested end to end.
  • The effective configuration matches the intended design and no unexplained error remains in the relevant logs.
  • Temporary troubleshooting controls have been removed and monitoring remains normal.
  • The result, evidence and any follow-up action are documented.

Rollback

  • Restore the configuration, policy, binding, route, credential assignment or service state recorded in the baseline when the change does not meet its success criteria.
  • Remove temporary rules, test objects and diagnostic settings that were introduced only for the procedure.
  • After rollback, repeat the minimum health checks to confirm that the previous service level has been restored.

Troubleshooting / common errors

  • A successful ping does not prove that the application protocol is allowed or listening.
  • When packet and policy evidence disagree, capture the same test flow at more than one point in the path.
  • If the result changes between tests, compare source, destination, identity, time and policy context before changing additional settings.
  • If a command succeeds but the application still fails, continue at the next protocol or application layer instead of widening access.
  • If the expected evidence is missing, verify that logging, auditing and the test path actually cover the failing component.
  • If the change does not improve the measured symptom, restore the previous state and reassess the working hypothesis.

Official and vendor references preserved from the original procedure

♡ 0