The Foundation for Your Hybrid Cloud
Recently active
i have two cluster main cluster and DR cluster , what is the best practice to upgrade two cluster , can i start with main or DR first , what about replication it is effected with upgade?
Hi community,I'm in the middle of deploying a 3-node Nutanix AHV cluster and have a question about live migration traffic behavior with and without the Backplane LAN configured.My current setup:- 3 Nutanix AHV nodes- 10G SFP+ uplinks for VM traffic (VLAN 1 - e.g 10.10.100.0/24)- 1G RJ45 uplinks for DMZ and Corp networks on separate VLANs- Backplane LAN not yet configuredMy questions:1. Without the Backplane LAN configured, when a host is placed into maintenance mode and VMs are live migrated to other nodes, does that migration traffic travel over the VM's own network? For example, if a VM is on my DMZ network (1G), does its live migration traffic also go over that 1G link?2. Once the Backplane LAN is configured on a dedicated VLAN (e.g. VLAN 10 - 192.168.10.0/24) on my 10G uplinks, will ALL live migration traffic route over the backplane regardless of which network the VM is connected to?3. Is there any official documentation that specifically covers this default migration traffic beh
I am setting up recovery plans for disaster recovery between our two clusters. I have been through the documentation multiple times, and I still am confused by the test failover/failback subnets that are configured as part of the recovery plan. Production is easy. We have the primary cluster in one location in Subnet A. We have a recovery cluster in another location in Subnet B. I just set the production subnets to be Subnet A and B at each location respectively with their matching gateway and prefix. The Test Failover/Failback subnets continue to confuse me. The documentation indicates they should be isolated and non-routable subnets which makes sense. But WHERE are those subnets defined? We use external IPAM. Each cluster is connected to a core switch stack at its location. We have a few different subnets/VLANs defined on those core switches, with a defined IP range, /24 in each case. All VMs have static IP addresses, though the core switches do run DHCP for each subnet. We also have
Good morning all, i have the problem, that one CVM out of a 3-Node-Cluster drives me crazy… For some reason, the CVM went unresponsive regarding it´s cluster-services… CVM itself is available via SSH and I´m able to login using the Nutanix-User CVM cluster status says: nutanix@NTNX-xxx-B-CVM:172.18.95.76:~$ cluster status | grep -v UP2026-05-08 04:50:48,386Z INFO MainThread zookeeper_session.py:226 Ignoring passed host_port_list: zk1:9876,zk2:9876,zk3:9876 because the passed host_port_list appears to have been copied from the environment variable or gflag.2026-05-08 04:50:48,387Z INFO MainThread zookeeper_session.py:296 cluster is attempting to connect to Zookeeper (unestablished session (object 0x7f6bd2b5d220)), host port list zk1:9876,zk2:9876,zk3:98762026-05-08 04:50:48,387Z INFO MainThread patterns.py:75 Creating a new instance for ZookeeperSession[('client_id', None), ('connection_timeout', None), ('host_port_list', 'zk1:9876,zk2:9876,zk3:9876'), ('use_zk_mt', None)]2026-05-08
I'm trying to deploy Prism Central (PC) from Prism Element (PE). Previously, I had registered a PC to the cluster and then unregistered it to redeploy a new one.However, during the new deployment, at the “Size and Scale” step, no selectable options (Small/Medium/Large) are shown — only the description text appears.Current situation:No dropdown or selectable options available No explicit error message Cluster is running normallyI'm not sure if this issue is related to the previous register/unregister of PC, or due to resource constraints, version mismatch, or a UI issue.Has anyone experienced this after unregistering PC? Is there any cleanup required on the cluster before redeploying Prism Central?
Hi Experts !We recently upgraded one of our Nutanix clusters to AOS 6.8.1.8 and started facing an issue with the Cerebro service shortly after.Symptoms observed:Remote replication alerts after the upgrade.Remote site configuration warning indicating that some CVM/SVM IPs were not properly configured on the peer site.Cerebro entering a crash loop?cluster status showing constantly changing/high PIDs for Cerebro.cerebro.FATAL reporting an error similar to: Check failed: citer != (pd.second)->snapshot_uuid_map().end() with a snapshot stuck in a pending action.It looks like there is a stale or orphaned snapshot operation in the WAL / metadata, likely triggered after the upgrade while replication configuration was not fully consistent.Has anyone already seen this behavior after an AOS upgrade?Did you resolve it by fixing the remote site configuration only, or did it require Nutanix Support intervention to skip/clean the offending WAL operation?Thanks :)
Hello I am deploying a single node cluster but I have the next errors during the creation:Node log:2026-04-07 07:10:22,518Z ERROR svm_rescue:329 Unable to create CVM UUID Marker. Error: [Errno 2] No such file or directory: '/sys/class/dmi/id/product_uuid'2026-04-07 07:10:22,518Z ERROR svm_rescue:1430 Unable to create CVM UUID marker.In the cluster log:2026-04-07 13:46:36,188Z flagvalues.py:537 ERROR Trying to access flag gateway_server_jsonrpc_url before flags were parsed.Traceback (most recent call last): File "gflags\flagvalues.py", line 535, in __getattr__gflags.exceptions.UnparsedFlagAccessError: Trying to access flag gateway_server_jsonrpc_url before flags were parsed.2026-04-07 13:46:36,189Z connectionpool.py:1001 DEBUG Starting new HTTPS connection (1): 10.25.0.103:22002026-04-07 13:46:36,646Z connectionpool.py:456 DEBUG https://10.25.0.103:2200 "POST /jsonrpc HTTP/1.1" 200 832026-04-07 13:46:36,686Z imaging_step_cluster_init.py:620 DEBUG Couldn't get status for services from g
how can i download nutanex AVH
Hello everyone,I am planning a migration from VMware to Nutanix AHV.The environment still includes some legacy operating systems, such as:Windows XP / Windows Server 2003 Windows Server 2008 (non-R2) Very old Linux distributions (kernel 2.6 era)I fully understand that:These OS versions are not supported on AHV Migration is technically difficult and not recommended This would be a best-effort / temporary solutionHowever, due to VMware licensing constraints, the customer must exit VMware and migrate these VMs anyway.My question to the community is simple:👉 In this kind of situation, what has been the most effective or realistic approach you have used to migrate legacy OS VMs from VMware to AHV?Any real-world experience or advice would be appreciated.Thank you.
I recently deployed Nutanix , but when I am opening web browser it is showing and I belive that it might be the certificate issue. I issued a commad to check the certificate in CVM and it is showing below and my system team already passed the date and time, what might be the other issue. Cluster is showing UP and healthy. upstream connect error or disconnect/reset before headers. retried and the latest reset reason: connection failure, transport failure reason: TLS error: 268435581:SSL routines:OPENSSL_internal:CERTIFICATE_VERIFY_FAILED
Are we able to configure a recovery plan in Nutanix Leap when we have Production and DR clusters (10 Clusters each) managed across two separate PCVMs (both running version 7.5.0.5)?Currently, while creating the recovery plan, I can select individual VMs, but I do not see an option to select at the cluster level. Is this expected behavior when using two PCVMs, and how can we properly configure a recovery plan in this setup?Any guidance on achieving this design would be helpful.
Hi I have a customer that is asking me for some feature on Nutanix (AHV) that may provide some kind of immutability for the VM snapshots, so in the event of an admin account hack the snapshots could not be destroyed.I was looking for something like that on AHV and I discovered Secure Snapshots with Approval Policies in Prism Central and I’d like to confirm a few points regarding requirements and licensing.https://portal.nutanix.com/page/documents/details?targetId=Disaster-Recovery-DRaaS-Guide-vpc_7_5:ecd-approval-policies-dr-pc-c.htmlContext:The customer has two AHV clusters, each one with his own Prism Central Licensing: NCI Pro + Advanced Replication add-on (Metro/Sync already in use) + NUS Pro Goal: prevent accidental or malicious deletion of snapshots/recovery points, not VMs themselves (it would be also great but I think it can’t be protected with aprobal policies)So from the documentation, I understand that Secure Snapshots allows attaching an Approval Policy to a Protection Poli
I migrated several Windows VMs using Nutanix Move.After the migration, I noticed that the network configurations of some VMs do not match the source settingsIPv6 protocol: Disable → Enabled I checked KB-10818, but it doesn't seem to apply to my case.(Move degraded?) Network location: Public → Private (or Private → Public) Windows Firewall: Disable → EnabledWhy are these settings changing during migration? Environment:AOS: 6.10.1(Source) → 6.10.1(Target) Hypervisor: AHV 20230302.103003(Source) → AHV 20230302.103003(Target) Guest OS: Windows Server 2019 Nutanix Move version: 6.0
Hello,Today I have delivered a three node Cluster to customer.Cluster was Set Up a few days ago in my Environment with a flat Switch.Bond mode for both vswitches is active-backup.Unfortunately Customer had prepared his ToR switches with LACP.All hosts and CVMs cannot communicate.What is the recommended procedure to fix ist?Change bond Mode host by host with manage_ovs OR Let network team configure the switches (disable LACP) and configure the clusternodes via Prism, after that let Network Team enable LACPKind regards,Thomas
Creating single node cluster since foundation is presenting an error:in the node log:2026-04-07 07:10:22,518Z ERROR svm_rescue:329 Unable to create CVM UUID Marker. Error: [Errno 2] No such file or directory: '/sys/class/dmi/id/product_uuid'2026-04-07 07:10:22,518Z ERROR svm_rescue:1430 Unable to create CVM UUID marker. in the cluster log:026-04-07 13:46:36,188Z flagvalues.py:537 ERROR Trying to access flag gateway_server_jsonrpc_url before flags were parsed.Traceback (most recent call last): File "gflags\flagvalues.py", line 535, in __getattr__gflags.exceptions.UnparsedFlagAccessError: Trying to access flag gateway_server_jsonrpc_url before flags were parsed.2026-04-07 13:46:36,189Z connectionpool.py:1001 DEBUG Starting new HTTPS connection (1): 10.25.0.103:22002026-04-07 13:46:36,646Z connectionpool.py:456 DEBUG https://10.25.0.103:2200 "POST /jsonrpc HTTP/1.1" 200 832026-04-07 13:46:36,686Z imaging_step_cluster_init.py:620 DEBUG Couldn't get status for services from genesis: {'stat
Hi,I’m looking to deploy the L4 load balancers for domain controller services LDAP(S), NTP, & DNS, you know they type of thing that gets addresses programmed into devices/applications.I can create a LB session for an IP and have it listen on multiple ports 53,123,636 but when I gets round to selecting the destination VM Nic's I can only specify a single port for the destination. Is there any way to balance multiple services from a set of hosts with a single VIP, or will I need to create an LB session and separate VIP for each of the services?Also some of the documentation i read said that It’s possible to use VM Categories to select the destination VM’s, but I don’t seem to have that option, which version do I need to be running to get that?
Hi,I am trying to deploy an AHV cluster using Nutanix Foundation VM against HPE DX380 G12 servers with iLO7, but the deployment consistently fails in Phase 1 during "Preparing installer image".The relevant error in the VM Foundation log is:Imaging requires Foundation HTTPS to be enabled on port 8001 for nodes that require HTTPS boot media.What I have verified so farOn the Foundation VM:firewalld is not running port 8000 is listening port 8001 is not listening at all curl -vk https://127.0.0.1:8001 returns connection refused curl -v https://127.0.0.1:8000 returns wrong version number, which seems to confirm that port 8000 is only serving HTTP, not HTTPS ss -lntp | egrep '8000|8001' only shows *:8000 nginx -T does not show any active listen 8001 ssl configurationThe Foundation process listening on 8000 is: /home/nutanix/foundation/lib/python/bin/python3.9 /home/nutanix/foundation/bin/foundation --foreground Deployment behavior All nodes fail with the same status:fatal: Preparing installe
I have a 1 node clusterOuter AHV - 10.1.10.103Outer CVM/Prism - 10.1.10.82Data Service ip - 10.1.10.95Network-0 having vlan id 1490 - gateway - 10.1.149.1, useable ips 10.1.149.2 to 10.1.149.61Prism - https://10.1.10.82:9440/-------------------------------------------------------------------created a 3 node nested clusterAHV VM1 - 10.1.149.51AHV VM2 - 10.1.149.53AHV VM3 - 10.1.149.55Data Service ip - 10.1.149.11 Inner Prism - https://10.1.149.52:9440/nested CVM1 - 10.1.149.52nested CVM1 - 10.1.149.54nested CVM1 - 10.1.149.56vlan1 having vlan id 1490 - gateway - 10.1.149.1, useable ips 10.1.149.2 to 10.1.149.61-------------------------------------------------------------------I am trying to create a Distributed FSVM on the nested cluster, Creating Distributed FSVM on inner cluster on vlan 1490 Client Network - 10.1.149.12,10.1.149.13,10.1.149.14 Storage Network - 10.1.149.15,10.1.149.16,10.1.149.17,10.1.149.118FSVM creation is failing in cluster initialisation with error as CVM not ab
Hello, I have recently deployed nutanix with Foundation, and every went well and I can access the PE from web browser. But when I go the Virtual Switch section, I cannot See vs0 in Switch Tab, not sure what is the issue. I go to Cluster Setting and Reboot the cluster, but it keep on processing and so on entering in maintenance mode. Can anyone help. Regards
Cloud-based disaster recovery with Nutanix MST (Multi-cloud Snapshot Technology) gives organizations a cost-effective way to replicate VM snapshots to object storage targets like AWS S3, Azure Blob, or Nutanix Objects and recover workloads when it matters most. Object storage is purpose-built for durability and geographic distribution, which makes it an ideal replication target. The trade-off has always been recovery speed: pulling data back from object storage to the AOS cluster is slower than reading from local NVMe or SAN-backed storage, and for large workloads that gap translated into hours of hydration time before a recovered VM could power on. For latency-sensitive workloads including clinical applications, financial systems, and operational databases, that recovery window was the one remaining challenge to address.Nutanix Instant Restore changes the equation for MST-protected workloads. Now generally available in AOS and Prism Central 7.5.1, it decouples recovery from object sto
We have created Auto shutdown playbook of AHV host with scheduled but facing issue, when Prism Central ON , these scheduled initiated before time and we not able to to access prism central , so disable playbook task. Prism element started then PC started and excuted these playbook schedule and shutdown 2 AHV from Prism cluster.
Dear Community members,We are using Cisco HX with following specifications and are interested to know for the proposed nutanix configurations, we need to know the pros and cons on the proposed specs and also would like to know or need to discuss on future expensions.HCI Configuration:CPU 59.73%120.42 GHz Consumed201.6 GHz Total96 Total Physical CoresMemory 87.09%1559 GiB Consumed1790 GiB TotalRaw Storage 75.57%13.48 TiB Consumed17.84 TiB TotalOther DetailsIOPS 4327Read/Write Ratio44/56Estimated Daily Usage28.25 GiB Proposed Specications: Part No Description Qty NX-8170-G10-6515P-CM NX-8170-G10, 1 Node; 2x Intel Xeon 6515P processor (2.3 GHz/ 16-core/ 150W, Granite Rapids SP) per node 3 C-MEM-32GB-6400-CM 32GB Memory Module (DDR5-6400 RDIMM) 72 C-NVM-7.68TB-AB1A-CM 7.68 TB NVMe SSD - PCIe Gen5 (U.2) 12 C-LOM-10G2D1BT-CM LOM Module: Broadcom 10GbE, 2-port, Base-T NIC (BCM 57416)
HiI have a customer with an old Nutanix AHV Cluster:Prism Central: pc.2023.4.0.2 AOS: 6.5.6.6 AHV: el7.nutanix.20220304.511 We wil deploy a new Nutanix AHV Cluster:Prism Central: 7.5.0.6 AOS: 7.5.0.6 AHV: 11.0.0.2 But the problem here is Veeam compatibility between enviroments. Right now the customer has two Veeam 12 apliances, the initial proposal was to eliminate one of them and use a single Veeam appliance with Veeam13.However when I was checking the compatibility matrix I can see the following info:https://helpcenter.veeam.com/docs/vbr/userguide/ahv_system_requirements.html?ver=13https://portal.nutanix.com/page/compatibility-interoperability-matrix/software?partnerName=Veeam&solutionType=Data%20Protection&componentVersion=all&hypervisor=AHV&validationType=all As far as I understand from the previous links if you want to use Veeam 13 you need at least:Prism Central → pc.2022.6 or above AOS → 6.5.6.6 or above AHV → 10.3.1.1 or above So checking the inventory from
Hello,I have some hands-on experience with Nutanix AHV. Currently, we are starting a project that involves a full migration of our Nutanix environment from on-premises infrastructure to Azure NC2.I have been searching for an official step-by-step guide or a recommended migration workflow, but so far I have not been able to find a clear document that explains the correct process for planning and executing this migration.Could anyone please share documentation, best practices, or a high-level migration flow for moving workloads from Nutanix AHV to Azure NC2?Any guidance or references would be greatly appreciated.Thank you.
Hello, I’ve deployed several test Nutanix Single-Node clusters for test purposes. Actually I’m encountering strange error during LCM update for AHV (I’m running version 10.0.1.6 update to version 10.3.1.2 is available). During update process LCM returns an error: Operation Failed. Reason: LCM operation update failed on leader, ip: [x.x.x.x] due to Upgrade failed: Unable to run upgrade plugin prepare_reimage: exit status 1Plugin stdout:Staging AHV media to /media/ahv.iso'/dev/sdd' -> '/media/ahv.iso'Preparing for reimagingAHV installation on LVM, PV device: /dev/sda3Plugin stderr:No symlink matching */md/* or */disk/by-id/* found for block device /dev/sdaFailed to prepare for reimaging. Logs have been collected and are available to download at /home/nutanix/data/logbay/bundles/auto-lcm-log-collection-2026-03-12-27553431714.zip I have no idea what can I do to resolve this issue, I’ve also tried deploying Prism Central but it seems that not a single version is compatible, because in
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.