VergeOS vSAN Scale Out Guide
Standard operating procedure for scaling out a VergeOS system by adding a new node, including preparation, pre-checks, execution, post-verification, and rollback procedures.
This guide provides best practices for safely scaling out a VergeOS system by adding a new node. It focuses on data security and system resilience through a methodical approach.
Prerequisites
Administrative access to VergeOS system
IPMI access to all nodes in the cluster
New node hardware properly racked and connected
Network connectivity verified for new node
USB drive with appropriate VergeOS installer version
Estimated time: 2-4 hours depending on data rebuild and cluster size
Important
Ensure you have a maintenance window that accounts for potential rebuild time. vSAN data redistribution can take several hours depending on the amount of data and storage performance.
Preparation Phase
Complete these steps well before your scheduled maintenance window:
System Readiness
Hardware Preparation
Documentation and Planning
Resource Verification
Pre-Scale Out Verification
Perform these checks on the day of scheduled maintenance, before beginning the scale out:
Current State Verification
Execution Phase
Adding the New Node
For detailed step-by-step instructions on performing the vSAN scale out, please refer to:
Key Process Overview
Boot new node from USB installer
Configure networking on the new node
Join node to cluster via VergeOS UI
Monitor vSAN redistribution - this is the longest phase
Wait for completion - vSAN tiers must return to green status
Critical Wait Period
During the scale out process, vSAN Tiers will show a yellow status during the rebuild stage. It is essential to wait for the vSAN tier to return to a "green" healthy status before continuing with any other operations or considering the scale out complete.
Monitoring Progress
You can monitor vSAN rebuild progress in the VergeOS UI under System -> vSAN. The rebuild process will show percentage completion and estimated time remaining.
Post-Scale Out Verification
After the scale out completes, verify the system is operating correctly:
System Health Checks
Performance Validation
Troubleshooting
Common Issues
vSAN tier remains yellow after extended time
Check network connectivity between nodes
Verify storage tier health in System -> vSAN
Ensure sufficient bandwidth between nodes
Contact support if rebuild stalls for more than expected timeframe
New node not appearing in WebUI
Verify network configuration and connectivity
Check DHCP/static IP assignment
Confirm USB installer version matches cluster version
Review node installation logs
Performance degradation during rebuild
This is normal during vSAN redistribution
Consider scheduling during low-usage periods
Monitor and adjust workload if necessary
Network fabric issues
Verify Core1 and Core2 network paths
Check physical network connections
Confirm VLAN configuration matches existing nodes
Rollback Procedure
Rollback Considerations
If issues occur, the safest approach is to pause, investigate the unexpected behavior, and then proceed again with a clear understanding of the problem. Complete node removal may require vSAN data redistribution.
If Rollback is Required
Stop the installation if still in progress
Power down the new node to prevent data corruption
Allow vSAN to stabilize if data was partially migrated
Remove node from cluster via VergeOS UI if it was successfully added
Monitor system stability before rescheduling the scale out
Next Steps
After successful scale out completion:
Emergency Support
Should you have any issues with your VergeOS scale out, please contact our support team for immediate assistance.
Last updated
Was this helpful?