Get guidance, share wins, and ensure smooth Nutanix deployments.
Recently active
Hi, I’m new to Nutanix, coming from the vSphere world.I am working on getting our new Nutanix infrastructure set up to run Qualys scans for security. I have been looking at the following document:https://portal.nutanix.com/page/documents/kbs/details?targetId=kA07V000000LXYqSAOI am getting stuck at determining which accounts to use to set up authentication, so that Qualys can run its scans on the Hypervisors, CVMs, and Prism Central appliances. My research shows that adding new accounts via the OS to these components is not supported. Does that mean that Qualys then has to log in as root, nutanix, and admin respectively, to scan these components? Or am I looking in the wrong place for the Qualys account to be set up, to scan for vulnerabilities and compliance? I feel like I'm missing something here. Thanks!
Hi allI have a cluster with vswitch active - activeI have new nodes to expand and I did theses steps below 1 - root@ahv# ./network_configuration to configure ips and vlan tag2 - Configure lacp 2.1 ssh root@192.168.5.1 "ovs-vsctl set port br0-up other_config:lacp-time=fast" 2.2 ssh ovs-vsctl set port br0-up other_config:lacp-fallback-ab=true 2.3 ssh root@192.168.5.1 "ovs-vsctl set port br0-up lacp=active" nutanix@CVM$ ssh 2.4 ssh ssh root@192.168.5.1 "ovs-vsctl set port br0-up bond_mode=balance-tcp"3 - Expand cluster The nodes were joined success in the cluster, but when I checked the vswitch configuration of theses nodes was changed to active - backupDid I do something wrong? What better way? Thank you
Below are new knowledge base articles published on the week of April 9-15, 2023.KB 12864 - Alert - A400117 - PolicyEngineVersionMismatch KB 13793 - NDB - scheduled daily snapshots break MSSQL differential backup chain KB 14359 - Prism Central - Redeploy errored out Infrastructure App from My Apps in Admin Center KB 14419 - Nutanix Objects Alert - kUnsynchronizedClock KB 14420 - Nutanix Objects Alert - kLowWALStorage KB 14421 - Nutanix Objects Alert - kMultipleFailedLeaderElections KB 14508 - No license eligible to convert KB 14517 - Unable to power on or migrate VMs to GPU enabled host KB 14599 - Prism Central 1-click deployment may fail when the gateway does not respond to ICMP Echo requests KB 14608 - Exporting an AHV VM using UEFI to OVA and importing to ESXi cluster causes an unbootable VM.Note: You may need to log in to the Support Portal to view some of these articles.
Below are new knowledge base articles published on the week of April 2-8, 2023.KB 13467 - No compatible products found when uploading cluster summary file KB 13903 - LCM Pre-check: "test_node_removal_status" KB 14104 - NCC Health Check: check_cvm_ssh_security KB 14121 - LCM Pre-check: "test_expand_cluster_status" KB 14277 - LCM Precheck: test_lacp_configuration KB 14304 - LCM 2.6: LCM upgrade will not fail when num_candidates > 0 KB 14407 - NDB | MSSQL AG Clone group deployment fails KB 14465 - Cluster outage can occur during multi-node removal when services go down on node which is yet to be detached KB 14522 - Slow start time of Siemens NX application on AHV due to FlexLM network query KB 14558 - Nutanix Cloud Cluster (NC2) on AWS - High space utilization during/post cluster recovery from S3 KB 14561 - License alert - Cluster is under-utilizing the Files license capacity. Usage: 0.00 TiB, Allowed license capacity: 1.0 TiB KB 14566 - NDB - OracleDB patching leaves old home and grid
Hello, I understand that if the node fails, the ha function is activated and attempts to recover. If one node fails in a Data Resilience Status Critical state due to a lack of storage space, will the cluster collapse? What happens.
I ma installing remotely Nutanix AHV cluster in very isolated environment, where for everything had to request firewall rules. IPMI is in different non-routable subnet, other that AHV and CVM. Firewall team confirmed that they don’t see any blocked traffic, but Foundation fails with error like this:2023-04-07 18:25:48,759Z WARNING Failed to register internal hypervisor: <Left ("Unable to satisfy any required capabilities for '<ExportedObjectDescriptor "command##.BindInternalHypervisor-1.0.0">'") at 0x52d10f8>2023-04-07 18:25:48,759Z INFO Tartarus initialization complete2023-04-07 18:25:48,996Z DEBUG Failed to load all plugins: Unable to satisfy required capability 'OPTIONAL_SMBIOS' for <ExportedObjectDescriptor "command##.LenovoRedfish_Pyghmi-1.0.0"> Skipping command##.GenericRedfish_Pyghmi-1.0.0, entity is disabled Unable to satisfy required capability 'OPTIONAL_SMBIOS' for <ExportedObjectDescriptor "command##.YadroRedfish_Pyghmi-1.0.0"> Unable to satisfy requi
Morning all!I have a 5 node cluster (5.20/ESXi7.0U3). its about 4-5 years old. Just got approval to upgrade and ordered a new 5 node cluster.Question: Is it best to expand current cluster and remove a node at a time, or set up net new and migrate?Looking for suggestions and experience.Thanks in advance!TJ
Can someone provide a brief explanation of Nutanix local key manager and how it handles the following in ‘???’:key access: access to the management console is restricted to authorized individuals abased on job function. changing / updating keys: encryption keys are updated from the management console, software encryption (rekey button) if necessary. revoking keys: ??? (when does revoking keys happen?) recovery keys: ??? (how do you recover if you have the key backup?) archiving keys: keys are archived (backed up) from the management console, managed keys where you download the key backup by setting up a recovery password to decrypt the backup file. activity logs: ??? (is there activity logs for keys? If yes, where is this stored and how long is the retention before the activity is overwritten?)BTW: I have this link already Native Local Key Manager (nutanix.com) but is does not have any details of the Thanks in advance.
Below are the top knowledge base articles for the month of March 2023.KB 1540 - [AOS Only] What to do when /home partition or /home/nutanix directory on a Controller VM (CVM) is full KB 7503 - NX Hardware [Memory] - DIMM Error handling and replacement policy KB 8885 - Alert - A15039 - IPMI SEL UECC Check KB 1381 - NCC Health Check: host_nic_error_check KB 1113 - HDD or SSD disk troubleshooting KB 4409 - LCM: Life Cycle Manager Troubleshooting Guide KB 2090 - AHV host networking KB 13476 - Domain Manager Internal Service has Stopped Working (helios) KB 3786 - Alert - A1081 - CuratorScanFailure KB 2473 - NCC Health Check: cvm_memory_usage_check KB 2475 - NCC Health Check: storage_pool_space_usage_check KB 5228 - NCC Health Check: pcvm_disk_usage_check KB 8514 - NCC Health Check: fs_inconsistency_check KB 13150 - NCC Health Check: cfs_fatal_check KB 13136 - NCC Health Check: microservice_infrastructure_status_check KB 13870 - Prism Virtual IP is configured but unreachable alert and VIP be
Below are new knowledge base articles published on the week of March 26-April 1, 2023.KB 13758 - NCC health check: usb_nic_status_check KB 14480 - LCM 2.5: Intel hardware DCF firmware upgrade failure since the KCS Policy Control Mode is currently set to "RESTRICTED" KB 14488 - Error: Protection domain 'XXXX' has entities in clone. Delete the clones before deleting this protection domain KB 14502 - ASMFD may cause instability on Linux VMs running on AHV KB 14504 - After enabling CMSP on PC the protected entities under remote AZ shows error "Failed to Fetch" KB 14510 - How to extend storage on SQL database servers managed by NDB KB 14530 - Nutanix Files: "Invalid start IP address" error during deployment KB 14532 - Unable to install NGT as CDROM is not recognized in a Windows VM KB 14538 - VMware Update Manager (VUM) service crashing caused by a large amount of connections made to vCenter KB 14564 - Powering on the VM on the AHV host may fail with the "Unexpected combination of memoryBac
Hi, How can I disable ani-affinity rules ?
The NX G6 demo equipment is currently updated with the latest firmware. If you try to install version 5.10 to test the AOS upgrade with the latest version of the foundation, the 4. foundation version will not mount, and the 5. latest foundation version will not configure CVM after rebooting after installation and installation will fail. Is it like this originally?
Below are new knowledge base articles published on the week of March 19-25, 2023.KB 13943 - NCC Health Check: mantle_keys_mirroring_check KB 14416 - HW : Sporadic Link down alerts in IPMI Health event logs on a G8 hardware KB 14491 - Nutanix Object - Limitations of an Object-enabled cluster KB 14495 - Genesis service keeps crashing if a public key with a leading space character added to the cluster in PRISM Cluster Lockdown Settings KB 14496 - Era DB-Server Agent upgrade operation failed. Reason: cannot import name 'convert_bytes_to_string' KB 14503 - Data Protection - Error: Specified protection domain XXXX is protecting a Vstore which has id yyyy KB 14512 - Cluster Expansion fails if the new node was previously a Compute-Only node with network segmentation enabled KB 14514 - HPE DX380 G10 Plus: Filesystem Errors and I/O errors on a hypervisor boot volume KB 14519 - AHV and ESXi Lenovo Whitley platform hosts updated to UEFI 1.31 may restart due to NMINote: You may need to log in to th
I have a hyper-v vm that I’m looking to move to AHV. it’s a database with iscsi mounts. I know in the past move was not able to convert them. Is that still the case? If so does anyone know a way to convert these datastores? thanks.
Hello,I would like to change the password expiration for a local user I have created in the prism element.I’ve tried to change it via the command:nutanix@cvm$ sudo chage -M <MAX-DAYS> admin I’ve managed to do that to the admin user, but when I change “admin” with my local user, It says that the user does not exist in etc/passwd. Is their any way of changing the password expiration for a local user that it is not “admin”.(The user is a cluster admin).
Below are new knowledge base articles published on the week of March 12-18, 2023.KB 13746 - NCC health check: LongRunningSubtasks KB 13913 - LCM inventory fails with the error "nothing to open" on pc.2021.9.0.5 and later due to frequent registration/unregistration KB 14088 - NCC Health Check: check_host_password_expiry KB 14307 - Unable to start docker due to inode exhaustion in Objects cluster. KB 14391 - AHV upgrade or host rolling reboot tasks may fail on AHV clusters with AMD CPU during test_cluster_config pre-check KB 14433 - Objects: Bucket deletion failed with "The bucket you tried to delete is not empty. You must delete all versions in the bucket." error KB 14450 - Oracle DB in Oracle Dataguard configuration shows high latency / lag while performing NDB operation KB 14467 - Cluster will fail to upgrade through LCM if AOS version is 5.10.x or 5.5.x KB 14477 - NCC check may fail to complete due to vCenter connectivity failure KB 14479 - AHV hosts or VMs may lose network connectiv
We are triying to install a new Nutanix Cluster with 2 nodes Lenovo HS1021. The nodes machine type is 7D20. When we launch the installation with the Foundation, the use ends in failed and we see the following errors:WARNING Failed to register internal hypervisor: <Left ("Unable to satisfy required capability 'HYPERV_VERSION' for <ExportedObjectDescriptor "command##.BindInternalHypervisor-1.0.0">") at 0x7efbc4d82bd8>2023-03-14 17:15:27,733Z WARNING Skipping <ImagingStepPreInstall(<NodeConfig(X.X.X.X) @35d0>) @80d0> because dependencies not met2023-03-14 17:15:27,731Z WARNING Skipping <ImagingStepRAIDCheckPhoenix(<NodeConfig(X.X.X.X) @35d0>) @d150> because dependencies not met, failed tasks: [<ImagingStepInitIPMI(<NodeConfig(X.X.X.X) @35d0>) @d390>]2023-03-14 17:15:27,736Z DEBUG Setting state of <ImagingStepPhoenix(<NodeConfig(X.X.X.X.) @35d0>) @67d0> from PENDING to NR2023-03-14 17:15:27,736Z WARNING Skipping <ImagingStepPho
Greetings, Does the Nutanix Local Key Manager (LKM) satisfy the recommendations/requirements to safely implement the Data at Rest Encryption?The documentation at: https://portal.nutanix.com/page/documents/details?targetId=Nutanix-Security-Guide-v6_5:wc-security-data-encryption-aos-wc-c.html has the warning: "Caution: DO NOT HOST A KEY MANAGEMENT SERVER VM ON THE ENCRYPTED CLUSTER THAT IS USING IT!! Doing so could result in complete data loss if there is a problem with the VM while it is hosted in that cluster." I too share this concern, which led me to investigate External Key Managers, but I am wondering how does using the LKM alleviate this risk? Also, as stated in the Nutanix Bible as well as here: https://portal.nutanix.com/page/documents/solutions/details?targetId=TN-2026-Information-Security:TN-2026-Information-Security "Now that Nutanix supports its own native LKM, Nutanix also takes the KEK and wraps it with a 256-bit encryption key called the machine encryption key (MEK). The
There are still many places where the old version is being used.Even if you try to install it with the previous version (AOS 5.10 or lower version) of the foundation on the G6 device to test the upgrade, it cannot be installed.Is it because all the firmware is up to date? It was the version that was originally installed. AOS 5.10 versions are not installed now?In foundation the log is "Device 2 :The Length of Name is not correct" The phrase continues to print. Please reply from those with experience.
Hi AllWhen I try to install AOS 6.5.2 I get the following errors. IOError: [Errno 2] No such file or directory: '/etc/nutanix/factory_config.json' 2023-03-07 13:19:54,652Z CRITICAL svm_rescue:926 No suitable SVM boot disk found. 2023-03-07 13:19:54,652Z INFO svm_rescue:114 exec_cmd: sync; sync; sync 2023-03-07 13:19:54,658Z INFO svm_rescue:114 exec_cmd: umount -R /mnt/disk 2023-03-07 13:19:54,663Z INFO svm_rescue:114 exec_cmd: umount -R /mnt/data] 2023-03-07 13:19:52,159Z INFO Imaging thread 'svm' failed with reason [None] 2023-03-07 13:19:52,164Z CRITICAL Imaging thread 'svm' failed with reason [None] 2023-03-07 13:19:52,200Z ERROR Exception in running <InstallHypervisorKVM(<NodeConfig(172.16.150.9) @b5d0>) @ee10> Traceback (most recent call last): File "foundation/imaging_step.py", line 161, in _run File "foundation/imaging_step_hypervisor.py", line 47, in run File "foundation/imaging_step.py", line 353, in wait_for_event StandardError: Received "fatal" in waiting for ev
Below are new knowledge base articles published on the week of March 5-11, 2023.KB 12446 - NCC Health Check: pc_to_ahv_secondary_ip_reachability_check KB 14162 - Foundation Failing Due to Incorrect Date or TimeNote: You may need to log in to the Support Portal to view some of these articles.
My 3 node cluster consiting of 1065-G5 nodes is coming EOL this year. I will be replacing the hardware with a NX-3360N-G8 cluster. What is the recommended workflow for replacing the hardware? One thing to note is the current cluster is licensed with Prism Pro and the new one will be licensed with Starter as I’m not using all the features in Prism Pro.Is the replacement as easy as expanding the cluster one node at a time, migrating the workloads and then removing the old cluster? Will there be any hiccups with difference in licensing?
Hey guysI am reinstalling three Lenovo HX5521 nodes and I got stuck on the below errorFoundation IP not set. Try running the “set_foundation_ip_address” script on the desktopI'm running the foundation vm on the same subnet of the IPMI interfaces connected directly to an unmanaged switch, without VLANs or any other configuration. connectivity is perfect. I already reviewed all the settings and tried to reimage foundation using ESXi and AHV. Both result in the same error.Im running;Foundation_VM-5.2.2 AOS euphrates-5.20.3 LTS VMware-ESXi-7.0.1 or AHV-20201105.2244I've been trying to solve this problem for three days now, but so far I haven't found any clues. Any help will be greatly appreciated. Thanks! 👷🏽
Below are new knowledge base articles published on the week of February 26-March 4, 2023.KB 13178 - NCC Health Check: cluster_node_count KB 13481 - Increasing /dev/sda3 filesystem partition size on Policy VM upgraded to 3.6.1 KB 13653 - Nutanix Database Service | Pulse Telemetry KB 14018 - Alert - A130365 - PauseStretchTriggeredByWitness KB 14032 - Upgrading to NDB 2.5.1 for HA-enabled setups fails KB 14303 - Unable to delete a Nutanix Storage Container KB 14326 - AOS upgrade to 6.1.x (or above) stuck on network segmentation enabled cluster with mixed hypervisors. KB 14366 - Nutanix Database Service - MongoDB brownfield to greenfield conversion fails after upgrade from a version lower than 2.5.1 to a higher one KB 14367 - IPMI Network Configuration Requirements & Best Practices KB 14372 - Data Lens - how to re-enable file server after been disabled KB 14373 - Storage Container Usage may increase after AOS upgrade in ESXi clusters with Thick Provisioned virtual disks larger than 4Ti
Hello,We have an unsupported configuration on 2016 stretch cluster: 3 NIC teaming with LACP → should be 1 NIC Teaming with a switch independent The goal is to upgrade to server 2022, we have the manual option from TAC to upgrade to 2019 than 2022.Note based on what we have seen we have the following points:-From the doc. https://portal.nutanix.com/page/documents/details?targetId=Acropolis-Upgrade-Guide-v6_6:upg-cluster-upgrade-recommend-hyperv-r.html Upgrade to Windows Server 2022 Hyper-V from an LACP enabled Hyper-V 2019 cluster is not supported. Enabling Link Aggregation Control Protocol (LACP) for your cluster deployment is supported when upgrading hypervisor hosts from Windows Server 2016 to 2019. On the other hand in doc. https://portal.nutanix.com/page/documents/kbs/details?targetId=kA00e000000LLI8CAO the LACP is not supported on 2016 and 2022, since we have seen the following error as seen below (the ISO image was 2019 for the upgrade, not 2022 we can not upgrade to 2022 direc
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.