Adding and Removing Hosts¶
Adding a Host to an Existing Cluster¶
Hosts can be added to an existing HVM cluster to increase compute capacity or improve failure tolerance.
Note
The agent quorum, Corosync, DLM, and GFS2 details on this page apply to layouts 1.3 and 2.0. See Legacy HVM Clusters for Legacy host management.
Prerequisites¶
Before adding a new host:
The new host meets all requirements listed in Building New HVM Clusters (OS, CPU, memory, network, storage)
The new host has network connectivity to all existing cluster hosts on port 7443
Shared storage (iSCSI targets) is accessible from the new host
Sufficient license capacity is available
Procedure¶
Navigate to
Infrastructure > Clusters > [Cluster]Click + Add Worker
Provide the new host’s SSH IP, hostname, and credentials
Click Complete
What Happens Automatically¶
When a host is added to a layout 1.3 or 2.0 cluster, HPE Morpheus Software performs the following automated operations:
Package installation and host preparation — The same provisioning phases run as during initial cluster creation (KVM, Corosync, DLM, layout-specific networking, firewall, libvirt)
Corosync authkey distribution — The cluster’s Corosync authentication key is retrieved from Cypher and installed on the new host
Corosync configuration update — The new host’s node block is added to
corosync.confon all existing hosts, theconfig_versionis incremented, andcorosync-cfgtool -Ris run on each host to reload the configuration live (no restart required)DLM join — The new host joins the DLM cluster and gains access to existing lockspaces
GFS2 journal addition — For each HPE Clustered Datastore (Shared LUN), a new journal is added using
gfs2_jaddif the current journal count is less than the new total node countiSCSI target discovery — All configured iSCSI targets are discovered and auto-logged-in on the new host
Quorum list update — HPE Morpheus Software sends updated
quorumInfoto all agents in the cluster (including the new host). Each agent persists the updated member list to.quorum-nodesand begins peer-to-peer quorum pinging with the new hostStorage pool creation — libvirt storage pools are created on the new host for each existing datastore
Quorum Impact¶
Adding a host changes the quorum majority calculation:
Previous Size |
New Size |
New Majority Required |
Failure Tolerance |
|---|---|---|---|
3 |
4 |
3 |
1 host |
4 |
5 |
3 |
2 hosts |
5 |
6 |
4 |
2 hosts |
6 |
7 |
4 |
3 hosts |
Note
Adding a 4th host to a 3-node cluster increases the majority requirement from 2 to 3 without improving failure tolerance. Consider adding hosts in pairs (3→5, 5→7) for optimal availability.
Verification After Adding¶
After the new host is provisioned:
Navigate to
Infrastructure > Clusters > [Cluster] > Summary > QuorumpanelConfirm the total node count has increased
Confirm the new host shows as ONLINE
On the new host, verify Corosync membership:
corosync-quorumtool -lVerify DLM sees all members:
dlm_tool status -vCompare physical-interface inventory with
sudo hvmcli interfaces list --filter ethernet. If an interface was added to the host after enrollment, follow the Netplan and UI refresh procedure in Preparing HVM Hosts before assigning it to a Virtual Switch.
Removing a Host from an Existing Cluster¶
Hosts can be removed from an HVM cluster when decommissioning hardware, reducing cluster size, or replacing failed nodes.
Warning
Removing a host reduces the cluster’s failure tolerance. Ensure the remaining cluster size still meets your availability requirements before proceeding.
Prerequisites¶
Before removing a host:
Confirm the cluster layout and ensure the resulting cluster will still have quorum (at minimum 3 nodes for a single-site cluster)
Confirm the Quorum panel reports ACHIEVED, all surviving hosts are reachable, lockspaces are healthy, and shared datastores are mounted with healthy paths
Evacuate all VMs from the host using maintenance mode (see Host Maintenance & Evacuation)
Verify no VMs are pinned to the host that cannot be moved
Confirm no local or raw-device VM remains on the host and that remaining hosts have the required VM networks and storage access
Procedure¶
Place the host in maintenance mode to evacuate VMs (see Host Maintenance & Evacuation)
Navigate to
Infrastructure > Clusters > [Cluster] > HostsSelect the host to remove
Click Remove
Confirm the removal
Wait for the removal operation to finish; do not manually edit Corosync or Agent quorum files
What Happens Automatically¶
When a host is removed from a layout 1.3 or 2.0 cluster, HPE Morpheus Software performs the following:
GFS2 unmount — All HPE Clustered Datastores (GFS2 filesystems) are unmounted on the departing host
Cluster services stop — DLM and Corosync are stopped and disabled on the departing host
Corosync configuration update — The departing host’s node block is removed from
corosync.confon all remaining hosts. The configuration is archived (backed up to/etc/corosync/archive/),config_versionis incremented, andcorosync-cfgtool -Rreloads the configuration live on each remaining hostDatastore cleanup — Each datastore’s location reference to the removed host is cleaned up
Quorum list update — Updated
quorumInfois sent to all remaining agents. The departing host is removed from.quorum-nodeson all cluster members
Quorum Impact When Removing¶
Previous Size |
New Size |
New Majority Required |
Failure Tolerance |
|---|---|---|---|
7 |
6 |
4 |
2 hosts |
6 |
5 |
3 |
2 hosts |
5 |
4 |
3 |
1 host |
4 |
3 |
2 |
1 host |
Important
Never reduce a single-site cluster below 3 nodes. A 2-node cluster cannot achieve quorum if either node fails, resulting in a complete cluster outage.
Handling Offline Host Removal¶
If the host to be removed is offline (powered off or unreachable):
First confirm the host is powered off or otherwise cannot access shared storage. If isolation cannot be proven, do not remove it; contact HPE Support
Use the same UI Remove action. On layout 1.3 and later, HPE Morpheus Software removes the node from the Corosync ring using surviving online hosts and updates Agent quorum membership
If the departing node is still active in the Corosync ring, or the updated Corosync configuration cannot be written consistently to all survivors, HPE Morpheus Software aborts the reload and raises an alarm rather than risk partitioning the ring
If no other hosts are online, cleanup cannot proceed and HPE Morpheus Software raises an Unable to remove host alarm
Do not publish or use direct host-shell removal commands as a substitute for this workflow. If the UI operation raises an alarm or cannot establish a safe ring departure, retain the host record and contact HPE Support.
Post-removal Verification¶
Confirm the host no longer appears in the cluster Hosts list
Confirm the Quorum panel reports ACHIEVED with the new member count and no removed host under down, fenced, or unclean members
Confirm Corosync and DLM list only the remaining members and lockspaces are healthy
Confirm every shared datastore is mounted and healthy on every remaining host
Confirm evacuated or failed-over VMs run once, on hosts with their required networks and storage