For the complete documentation index, see llms.txt. This page is also available as Markdown.

VergeOS vSAN Scale Out Guide

Standard operating procedure for scaling out a VergeOS system by adding a new node, including preparation, pre-checks, execution, post-verification, and rollback procedures.

This guide provides best practices for safely scaling out a VergeOS system by adding a new node. It focuses on data security and system resilience through a methodical approach.

Overview

Scaling out a VergeOS system requires careful planning and execution to ensure minimal disruption to your services. This guide breaks down the process into main phases:

  1. Preparation - Steps to take before your scheduled maintenance window

  2. Pre-Scale Out Verification - Final checks immediately before beginning the scale out

  3. Scale Out Execution - The process of adding a new node

  4. Post-Scale Out Verification - Ensuring the scale out was successful

Environment-Specific Requirements

This guide covers general best practices. You may need to adapt these steps for your specific environment and requirements.

Prerequisites

  • Administrative access to VergeOS system

  • IPMI access to all nodes in the cluster

  • New node hardware properly racked and connected

  • Network connectivity verified for new node

  • USB drive with appropriate VergeOS installer version

  • Estimated time: 2-4 hours depending on data rebuild and cluster size

Preparation Phase

Complete these steps well before your scheduled maintenance window:

System Readiness

Use Case Examples

VDI may have several distinct images or resource pools. VPS may focus more on network functionality and connectivity.

Hardware Preparation

Documentation and Planning

Resource Verification

Pre-Scale Out Verification

Perform these checks on the day of scheduled maintenance, before beginning the scale out:

Current State Verification

Execution Phase

Adding the New Node

For detailed step-by-step instructions on performing the vSAN scale out, please refer to:

Key Process Overview

  1. Boot new node from USB installer

  2. Configure networking on the new node

  3. Join node to cluster via VergeOS UI

  4. Monitor vSAN redistribution - this is the longest phase

  5. Wait for completion - vSAN tiers must return to green status

Post-Scale Out Verification

After the scale out completes, verify the system is operating correctly:

System Health Checks

Performance Validation

Troubleshooting

Common Issues

vSAN tier remains yellow after extended time

  • Check network connectivity between nodes

  • Verify storage tier health in System -> vSAN

  • Ensure sufficient bandwidth between nodes

  • Contact support if rebuild stalls for more than expected timeframe

New node not appearing in WebUI

  • Verify network configuration and connectivity

  • Check DHCP/static IP assignment

  • Confirm USB installer version matches cluster version

  • Review node installation logs

Performance degradation during rebuild

  • This is normal during vSAN redistribution

  • Consider scheduling during low-usage periods

  • Monitor and adjust workload if necessary

Network fabric issues

  • Verify Core1 and Core2 network paths

  • Check physical network connections

  • Confirm VLAN configuration matches existing nodes

Rollback Procedure

If Rollback is Required

  1. Stop the installation if still in progress

  2. Power down the new node to prevent data corruption

  3. Allow vSAN to stabilize if data was partially migrated

  4. Remove node from cluster via VergeOS UI if it was successfully added

  5. Monitor system stability before rescheduling the scale out

Next Steps

After successful scale out completion:

Last updated

Was this helpful?