Get guidance, share wins, and ensure smooth Nutanix deployments.
Recently active
Using phoenix to reimage a node after a failed satadom, after phoenix is loaded, I uploaded AHV ISO so it can proceed to install AHV/AOS. It turns out I uploaded the lcm_ahv_el7.nutanix.20201105.2267.tar.gz instead of the AHV-DVD-x86_64-el7.nutanix.20201105.2267.iso. The install is now hung at 75% Node discovery succeeded. I have tried to install the proper iso, but it immediately failed at that same progress. How can I cancel this install so I can try again?I did find this error in the foundation log. It looks like the error happens if you use the html 5 client to mount the phoenix iso, but I am using the java client.
I am able to create a Move plan using the API and Ansible for migrating a VM from vSphere to Nutanix. However, I am unable to specify the guest credentials for the guest operations (like removing VMware Tools from the replicated vm during the migration) in the actual rest call (see code below) as is possible using the GUI. I cannot find any way in the API documentation on how to get the credentials into the request. Hope somebody can assist.- name: "Create Move Plan: Task 1.2a - Create Move Plan for OTA" uri: url: "https://{{ move_server_ota }}/move/{{ mapi_version }}/plans" method: POST validate_certs: no force_basic_auth: yes headers: Authorization: "{{ move_token_ota }}" body_format: json body: | { "Spec": { "Name": "{{ move_host_name | lower }}", "NetworkMappings": [ { "SourceNetworkID": "{{ move_source_network_id[0] }}", "TargetNetworkID": "{{ move_ota_target_network_uuid}}",
We have our Nutanix 5 node cluster installed with VMware vSphere 7.0U3. RF=2 so I should be bale to withstand a 1 node failure. I would like to test the HA capability of VMware/Nutanix by taking a node offline abruptly and ensure everything works as expected before placing production workloads on it.I was thinking of placing a test VM on a node and then pulling the power cables and ensuring that VM is powered up on a remaining host.Does anyone have any better or preferred way to test HA?
Below are the top knowledge base articles for the month of February 2023.KB 1540 - [AOS Only] What to do when /home partition or /home/nutanix directory on a Controller VM (CVM) is full KB 7503 - NX Hardware [Memory] - DIMM Error handling and replacement policy KB 8885 - Alert - A15039 - IPMI SEL UECC Check KB 1381 - NCC Health Check: host_nic_error_check KB 4409 - LCM: Life Cycle Manager Troubleshooting Guide KB 3827 - Alert - A130087 - Node Degraded KB 1113 - HDD or SSD disk troubleshooting KB 2090 - AHV host networking KB 2473 - NCC Health Check: cvm_memory_usage_check KB 13870 - Prism Virtual IP is configured but unreachable alert and VIP becomes permanently unreachable observed on AOS 6.5.1 and later KB 3786 - Alert - A1081 - CuratorScanFailure KB 8514 - NCC Health Check: fs_inconsistency_check KB 4519 - NCC Health Check: check_ntp KB 13150 - NCC Health Check: cfs_fatal_check KB 5228 - NCC Health Check: pcvm_disk_usage_check KB 8094 - NCC Health Check: disk_status_check KB 4158 -
Below are new knowledge base articles published on the week of February 19-25, 2023.KB 13146 - Flow Network Security may block traffic using virtual IP address or forwarded traffic KB 13587 - List of IAM Migration Alerts during Upgrade to Prism Central pc.2022.9 KB 14017 - Alert - A130354 - FailoverTriggeredByWitness KB 14308 - Network performance of VXLAN interfaces inside VMs severely degraded if VM is running on AHV host with Broadcom BCM57414 NICs with Hardware Generic Receive Offload (GRO) feature enabled KB 14332 - Metric IO stats missing from Prism or the Insights Portal KB 14343 - Prism Central - Templates view is empty for LDAP Users on Prism Central with enabled Microservices Platform (CMSP) KB 14344 - Prism Central: report is missing values or shows only a dash (-) KB 14351 - How to obtain quota limits via API after enabling Calm Policy Engine KB 14353 - Unable to change the VPN route priority in Nutanix DRaaS KB 14369 - AHV Metro/SyncRep : Services on Pacemaker unable to ru
Hello, i have nutanix 3060g8 i have error:Committed memory update intent that is stuck on 50% i tried to reboot cvm and all nodes but i still have same problem.cvm memory are 64gb i tired:ecli task.list include_completed=noTask UUID Parent Task UUID Component Sequence-id Type Statusf0ffcb92-96b1-426f-51d7-204729b269fe kGenesis 1 kCvmreconfig kRunning progress_monitor_cli --entity_id="f0ffcb92-96b1-426f-51d7-204729b269fe" --deleteany suggest ?
Hi Team,Which Firewall ports will be required for new node we are adding node in the existing cluster.There is VMware cluster with Nutanix cluster.kindly guide and suggest.
Hello eveyone,following my last post I eventually got my C node to be alive. (it was not booting nor being visible whatever I tries, always down).I can ping it with both IPv4 and IPv6 addresses, connect to it in SSH, run commands etc.Now my situation is : while I was struggling with my node down, I removed it from the cluster with Prism. I think I could bring it back after… But it does not work.As you can see here, the C node (#3) is now missing :I use the “Expand cluster” tools in Prism Element, I add the node manually, it is detectedI check the mode, validate and it starts to expand.But after a few minutes I get the error :Failure in pre expand-cluster tests. Errors: Failed to get HCI node info using discovery It seems like the cluster refuse to consider the node as “free”, or the node itself refuse to join because it thinks that it is still in the cluster.Thank you very much for any help you could provide
I got existing 3 nodes with Dell XC740 running with AHV. And I’ve order new 3 nodes x Dell XC750 but it came with factory install ESXi 7.0.Question is any special instruction do before adding new node to the existing cluster and I would like to new node XC750 to run the same AHV hypervisor in existing cluster too. So, can I follow expand step here Prism 6.5 - Expanding a Cluster (nutanix.com)?Thank you.
Below are new knowledge base articles published on the week of February 12-18, 2023.KB 14236 - PSOD on Ice Lake processor platforms after upgrading to ESXi 7.0 U3i due to an issue with microcode 0x0d000375 KB 14297 - Security Dashboard throws the error "Enable microservices infrastructure with internet connectivity for Security Dashboard to work" KB 14309 - NDB - MySQL software profile creation fails for commercial MySQL versions KB 14310 - NDB - Failed to create the software profile with error message: "local variable \'db_version\' referenced before assignment" KB 14318 - User defined Life cycle rule does not work in object bucketsNote: You may need to log in to the Support Portal to view some of these articles.
Whe doing an AHV update I noticed that when going into or out of Maintenance mode some of the VMs show as VM and some have their actual name - why is this?
I’d like to sort my powered off VMs under AHV by last date they were powered on?Can you share your thoughts on how to display through AHV?
Hello All,I tried to install AOS 6.5.x with foundation 5.3.2 and nutanix but I faced from some issues. please find the below logs for reference2023-02-16 09:10:32,277Z INFO [1081/2430] Hypervisor installation in progress2023-02-16 09:11:02,318Z INFO [1111/2430] Hypervisor installation in progress2023-02-16 09:11:32,358Z INFO [1141/2430] Hypervisor installation in progress2023-02-16 09:12:02,371Z INFO [1171/2430] Hypervisor installation in progress2023-02-16 09:12:32,413Z INFO [1201/2430] Hypervisor installation in progress2023-02-16 09:12:32,595Z WARNING Hypervisor installation takes longer than usual2023-02-16 09:13:02,453Z INFO [1231/2430] Hypervisor installation in progress2023-02-16 09:13:32,486Z INFO [1261/2430] Hypervisor installation in progress2023-02-16 09:14:02,525Z INFO [1291/2430] Hypervisor installation in progress2023-02-16 09:14:32,563Z INFO [1321/2430] Hypervisor installation in progress
I was intially told I could use Move to take VMs from old Cluster to New Cluster but alas Move does not move from AHV to AHV, this I find very bizarre as that should be more simple than moving esxi to ahv!So, I have been told to use Data Protection failover - this seems complicated and needs VMs to be stopped and restarted - I came across Live Migration using Leap - ah ha that looks like a good option but still apears complex. Do I really need 2 Prism Central instances? Could I add the second cluster to my exisiting Prism Central then use the Migrate function to move some VMs or is the limitation of different harware still the gotcha?Does anyone have any ideas how I can do this “easily” and “quickly” my boss ain’t a happy man as information from Nutanix has been misleading or downright incorrect.Any help appreciated.
I am curious why when sizing a Splunk workload and increasing the performance profile for the indexers it doesn’t increase the ingest rate per indexer rather than keeping it at 100 GB. By not doing this, it keeps the same number of indexers Nutanix thinks it needs even though each indexer has more CPU/RAM. I could be missing something simple here, but I would like to hear other thoughts.Capacity Planning Manual: Summary of performance recommendations
Hello, after a firmware upgrade, one host is locked DOWN in maintenance mode :CVM: 192.168.131.132 DownI can run a command to exit maintenance mode but it is not working and it is "Removed from metadata store" :nutanix@:~$ ncli host edit id=7 enable-maintenance-mode=falseId : …Hypervisor Address : 192.168.131.122Host Status : NORMALOplog Disk Size : 394 GiB (423,054,278,649 bytes) (3.9%)Under Maintenance Mode : false (ncli_manual)Metadata store status : Node is removed from metadata store...So I tried to recover it but the script fails :nutanix@:~$ python /home/nutanix/cluster/bin/lcm/lcm_node_recovery.py 192.168.131.122Recovering node 192.168.131.122Checking if the node 192.168.131.122 is in phoenixCurrent node status host Node 192.168.131.122 out of phoenix modeBringing host None out of maintenance mode Successfully put host None out of maintenance mode Bringing CVM 192.168.131.122 out of maintenance modeTraceback (most recent call last): File "/home/nutanix/cluster/bin/lcm/lcm_node_
Hi there, I need some explanation about fault tolerance in a particular situation.The configuration is one 4-nodes block with each nodes configured with 2 SSDs and 4 HDDs. Also Fault Tolerance FT=1 is configured.In that case, how many disk/node are protected ?
Below are new knowledge base articles published on the week of February 5-11, 2023.KB 13748 - Accessing the IPMI on-site KB 13807 - NCC Health Check: witness_fault_domain_check KB 14190 - SQL Server Standard Edition is not able to utilize all the CPUs assigned to the VM KB 14218 - Nutanix Database Service - VM provisioning fails if AD gMSA whitelist is used KB 14229 - After upgrading to pc.2022.9, issue with upgrade of microservice infrastructure OR MSP base services KB 14233 - Enabling Distributed Autorid on systems upgraded to Nutanix Files 4.2 or later KB 14243 - Nutanix Database Service operations fail with invalid authentication credentials KB 14278 - Alert - A20032 - Containers are marked for removal KB 14284 - NDB - Clone Database Failed KB 14295 - NDB Postgres DB Server provisioning fails with "Failed to configure storage for database" with 1Gb memory profile KB 14306 - After upgrading Prism Central from pc.2022.6 to pc.2022.9, the Security Dashboard might fail to runNote: You
Hi!We are running about 100 vm’s in our cluster - so find out a specific information manually for each vm is pretty hard.I want to get the information which vm’s are stored in a specific storage container with acli on cvm.I know that i could perform acli vm.get <vm-name> and look for the source_nfs_path of the vmdisks but it is way too much effort for 100 vm’s. Another way would be establishing a connection via SSH to /storage-container/.acropolis/vmdisk of of the cvm but i only see the vmdisk-uuids there, not the vm names i need.Is there any possibility, maybe with a for loop on each vm and a grep filter?
Hello Team,I hope you are doing well,I am learning more about Nutanix platform , can please share document(s) or KB(s) on how the physical disks are used in a Nutanix host? ie if the disk in slot 1 is reserved for CVM or m.2 is reserved to install the OS,I want to have my own guide that shows all models with their own disk layout according to their possible combinations -single SSD , dual or quad SSDs the rest of the slots are populated with HDDs -hybrid-or all slots are populated with SSDs ..ect.BRAdel
Below are new knowledge base articles published on the week of January 29-February 4, 2023.KB 12456 - NCC Health Check: ahv_gateway_check KB 13445 - Alert - A130377 - StaleVolumeGroup KB 13611 - Alert - A130382 - SynchronousReplicationPausedOnVM KB 13739 - Alert - A300441 - RecoveryPlanDynamicSubnetExtensionDeletionFailure KB 13773 - Alert-A130385 - SyncedVolumeGroupRunningAtSubOptimalPerformance KB 13796 - Alert - A130386 - AutoPlannedFailoverOfVolumeGroup KB 14126 - LCM Pre-check: test_ptagent_status KB 14173 - The cluster raises "A130184: Data At Rest Encryption key backup warning: Last user backup taken at MMM DD" even after performing the backup. KB 14211 - OpenShift 4.12 installation via the “openshift-install create install-config” command fails KB 14222 - Nutanix Database Service - Ubuntu/Debian DB Provisioning is failing due to static IP allocation failure KB 14224 - LCM UI is not available and inventory task stops at 0% KB 14234 - Alert - A160054 - File Server Partner Server
Below are the top knowledge base articles for the month of January 2023.KB 1540 - [AOS Only] What to do when /home partition or /home/nutanix directory on a Controller VM (CVM) is full KB 7503 - NX Hardware [Memory] - DIMM Error handling and replacement policy KB 8885 - Alert - A15039 - IPMI SEL UECC Check KB 2473 - NCC Health Check: cvm_memory_usage_check KB 13870 - Prism Virtual IP is configured but unreachable alert and VIP becomes permanently unreachable observed on AOS 6.5.1 and later KB 4409 - LCM: Life Cycle Manager Troubleshooting Guide KB 1113 - HDD or SSD disk troubleshooting KB 2090 - AHV host networking KB 3827 - Alert - A130087 - Node Degraded KB 13150 - NCC Health Check: cfs_fatal_check KB 4519 - NCC Health Check: check_ntp KB 8514 - NCC Health Check: fs_inconsistency_check KB 5228 - NCC Health Check: pcvm_disk_usage_check KB 3786 - Alert - A1081 - CuratorScanFailure KB 6153 - NCC Health Check: default_password_check and pc_default_password_check KB 4141 - Alert - A1046 -
Below are new knowledge base articles published on the week of January 22-28, 2023.KB 12548 - NCC Health Check: attached_vm_volume_different_recovery_clusters_check KB 12922 - Alert - A130369 - DistributedOplogFeatureDisabled KB 13040 - NCC Health Check: attached_vm_vg_protection_check KB 13100 - Alert - A130371/A130372/A130373/A130374/A130375 - DrNetworkSegmentationRemote KB 13513 - Alert - A160157 - FileServerVMcoreDetected KB 13889 - CRUD operations are not permitted on a stale Volume Group KB 13956 - CMSP enabled PC upgrade fails with error "Max retries done: unable to push docker image msp-registry.cmsppc.nutanix.com:5001/alpine:3.12" KB 13983 - Allow Out-of-band (OOB) snapshot for 3rd patry backups on RF1 VM KB 14177 - ESXi upgrade stuck at 14% due to vMotion taking longer than 15 minutes KB 14186 - Nutanix Files - Manual DNS verification fails even if the DNS entries are properly created KB 14195 - Acropolis service may crash with "Unexpected CAS error while trying to delete an
Below are new knowledge base articles published on the week of January 15-21, 2023.KB 10596 - Prism Central Pre-Upgrade Check: test_prism_central_cmsp_enablement_check KB 14074 - LCM Pre-check: test_empty_ilo_task_queue KB 14105 - After AOS upgrade to 6.x, users with the Prism Element Backup Admin role are unable to use AD credentials to log in to Prism Element KB 14167 - AHV upgrade is blocked in LCM 2.5.0.4 or newer if running vGPUs VMs are found KB 14185 - Power state of VMs managed by the Nutanix AHV Plugin for Citrix on AHV clusters changed to "Unknown" after upgrading AOS cluster to version AOS 6.6 or newer KB 14194 - Metrics-enabled (vhostmd) VMs fail to start on AHV 20220304.x KB 14198 - UpdateVmDbState task can often be seen on AHV clustersNote: You may need to log in to the Support Portal to view some of these articles.
Below are new knowledge base articles published on the week of January 8-14, 2023.KB 12052 - NCC Health Check: Prism Central disaster recovery PE-PC version compatibility check KB 13210 - NCC Health Check: pcvm_reboot_check KB 13272 - Pre-Upgrade Check: test_max_pe_nodes_registered_to_pc KB 13301 - Pre-Upgrade Check: test_prism_central_run_cmsp_preupgrade_check KB 13666 - Pre-Upgrade Check: test_prism_central_pcdr_compatibility_check KB 13676 - NCC Health Check: category_threshold_alert KB 13766 - Prism Central Pre-check: test_event_audit_counts_under_threshold KB 13872 - Pre-upgrade check: test_prism_central_pcdrv1_enabled_check KB 13979 - Custer creation fails on HPE DX G10 Plus nodes if there are no disks installed in a storage controller KB 14076 - In microservice infrastructure enabled Prism Centrals, /home directory space can be exhausted due to docker folders KB 14130 - LCM is running in Offline mode KB 14142 - Nutanix Kubernetes Engine - Pod to service communication does not wo
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.