Get guidance, share wins, and ensure smooth Nutanix deployments.
Recently active
I’d like to sort my powered off VMs under AHV by last date they were powered on?Can you share your thoughts on how to display through AHV?
Hello All,I tried to install AOS 6.5.x with foundation 5.3.2 and nutanix but I faced from some issues. please find the below logs for reference2023-02-16 09:10:32,277Z INFO [1081/2430] Hypervisor installation in progress2023-02-16 09:11:02,318Z INFO [1111/2430] Hypervisor installation in progress2023-02-16 09:11:32,358Z INFO [1141/2430] Hypervisor installation in progress2023-02-16 09:12:02,371Z INFO [1171/2430] Hypervisor installation in progress2023-02-16 09:12:32,413Z INFO [1201/2430] Hypervisor installation in progress2023-02-16 09:12:32,595Z WARNING Hypervisor installation takes longer than usual2023-02-16 09:13:02,453Z INFO [1231/2430] Hypervisor installation in progress2023-02-16 09:13:32,486Z INFO [1261/2430] Hypervisor installation in progress2023-02-16 09:14:02,525Z INFO [1291/2430] Hypervisor installation in progress2023-02-16 09:14:32,563Z INFO [1321/2430] Hypervisor installation in progress
I was intially told I could use Move to take VMs from old Cluster to New Cluster but alas Move does not move from AHV to AHV, this I find very bizarre as that should be more simple than moving esxi to ahv!So, I have been told to use Data Protection failover - this seems complicated and needs VMs to be stopped and restarted - I came across Live Migration using Leap - ah ha that looks like a good option but still apears complex. Do I really need 2 Prism Central instances? Could I add the second cluster to my exisiting Prism Central then use the Migrate function to move some VMs or is the limitation of different harware still the gotcha?Does anyone have any ideas how I can do this “easily” and “quickly” my boss ain’t a happy man as information from Nutanix has been misleading or downright incorrect.Any help appreciated.
I am curious why when sizing a Splunk workload and increasing the performance profile for the indexers it doesn’t increase the ingest rate per indexer rather than keeping it at 100 GB. By not doing this, it keeps the same number of indexers Nutanix thinks it needs even though each indexer has more CPU/RAM. I could be missing something simple here, but I would like to hear other thoughts.Capacity Planning Manual: Summary of performance recommendations
Hello, after a firmware upgrade, one host is locked DOWN in maintenance mode :CVM: 192.168.131.132 DownI can run a command to exit maintenance mode but it is not working and it is "Removed from metadata store" :nutanix@:~$ ncli host edit id=7 enable-maintenance-mode=falseId : …Hypervisor Address : 192.168.131.122Host Status : NORMALOplog Disk Size : 394 GiB (423,054,278,649 bytes) (3.9%)Under Maintenance Mode : false (ncli_manual)Metadata store status : Node is removed from metadata store...So I tried to recover it but the script fails :nutanix@:~$ python /home/nutanix/cluster/bin/lcm/lcm_node_recovery.py 192.168.131.122Recovering node 192.168.131.122Checking if the node 192.168.131.122 is in phoenixCurrent node status host Node 192.168.131.122 out of phoenix modeBringing host None out of maintenance mode Successfully put host None out of maintenance mode Bringing CVM 192.168.131.122 out of maintenance modeTraceback (most recent call last): File "/home/nutanix/cluster/bin/lcm/lcm_node_
Hi there, I need some explanation about fault tolerance in a particular situation.The configuration is one 4-nodes block with each nodes configured with 2 SSDs and 4 HDDs. Also Fault Tolerance FT=1 is configured.In that case, how many disk/node are protected ?
Below are new knowledge base articles published on the week of February 5-11, 2023.KB 13748 - Accessing the IPMI on-site KB 13807 - NCC Health Check: witness_fault_domain_check KB 14190 - SQL Server Standard Edition is not able to utilize all the CPUs assigned to the VM KB 14218 - Nutanix Database Service - VM provisioning fails if AD gMSA whitelist is used KB 14229 - After upgrading to pc.2022.9, issue with upgrade of microservice infrastructure OR MSP base services KB 14233 - Enabling Distributed Autorid on systems upgraded to Nutanix Files 4.2 or later KB 14243 - Nutanix Database Service operations fail with invalid authentication credentials KB 14278 - Alert - A20032 - Containers are marked for removal KB 14284 - NDB - Clone Database Failed KB 14295 - NDB Postgres DB Server provisioning fails with "Failed to configure storage for database" with 1Gb memory profile KB 14306 - After upgrading Prism Central from pc.2022.6 to pc.2022.9, the Security Dashboard might fail to runNote: You
Hi!We are running about 100 vm’s in our cluster - so find out a specific information manually for each vm is pretty hard.I want to get the information which vm’s are stored in a specific storage container with acli on cvm.I know that i could perform acli vm.get <vm-name> and look for the source_nfs_path of the vmdisks but it is way too much effort for 100 vm’s. Another way would be establishing a connection via SSH to /storage-container/.acropolis/vmdisk of of the cvm but i only see the vmdisk-uuids there, not the vm names i need.Is there any possibility, maybe with a for loop on each vm and a grep filter?
Hello Team,I hope you are doing well,I am learning more about Nutanix platform , can please share document(s) or KB(s) on how the physical disks are used in a Nutanix host? ie if the disk in slot 1 is reserved for CVM or m.2 is reserved to install the OS,I want to have my own guide that shows all models with their own disk layout according to their possible combinations -single SSD , dual or quad SSDs the rest of the slots are populated with HDDs -hybrid-or all slots are populated with SSDs ..ect.BRAdel
Below are new knowledge base articles published on the week of January 29-February 4, 2023.KB 12456 - NCC Health Check: ahv_gateway_check KB 13445 - Alert - A130377 - StaleVolumeGroup KB 13611 - Alert - A130382 - SynchronousReplicationPausedOnVM KB 13739 - Alert - A300441 - RecoveryPlanDynamicSubnetExtensionDeletionFailure KB 13773 - Alert-A130385 - SyncedVolumeGroupRunningAtSubOptimalPerformance KB 13796 - Alert - A130386 - AutoPlannedFailoverOfVolumeGroup KB 14126 - LCM Pre-check: test_ptagent_status KB 14173 - The cluster raises "A130184: Data At Rest Encryption key backup warning: Last user backup taken at MMM DD" even after performing the backup. KB 14211 - OpenShift 4.12 installation via the “openshift-install create install-config” command fails KB 14222 - Nutanix Database Service - Ubuntu/Debian DB Provisioning is failing due to static IP allocation failure KB 14224 - LCM UI is not available and inventory task stops at 0% KB 14234 - Alert - A160054 - File Server Partner Server
Below are the top knowledge base articles for the month of January 2023.KB 1540 - [AOS Only] What to do when /home partition or /home/nutanix directory on a Controller VM (CVM) is full KB 7503 - NX Hardware [Memory] - DIMM Error handling and replacement policy KB 8885 - Alert - A15039 - IPMI SEL UECC Check KB 2473 - NCC Health Check: cvm_memory_usage_check KB 13870 - Prism Virtual IP is configured but unreachable alert and VIP becomes permanently unreachable observed on AOS 6.5.1 and later KB 4409 - LCM: Life Cycle Manager Troubleshooting Guide KB 1113 - HDD or SSD disk troubleshooting KB 2090 - AHV host networking KB 3827 - Alert - A130087 - Node Degraded KB 13150 - NCC Health Check: cfs_fatal_check KB 4519 - NCC Health Check: check_ntp KB 8514 - NCC Health Check: fs_inconsistency_check KB 5228 - NCC Health Check: pcvm_disk_usage_check KB 3786 - Alert - A1081 - CuratorScanFailure KB 6153 - NCC Health Check: default_password_check and pc_default_password_check KB 4141 - Alert - A1046 -
Below are new knowledge base articles published on the week of January 22-28, 2023.KB 12548 - NCC Health Check: attached_vm_volume_different_recovery_clusters_check KB 12922 - Alert - A130369 - DistributedOplogFeatureDisabled KB 13040 - NCC Health Check: attached_vm_vg_protection_check KB 13100 - Alert - A130371/A130372/A130373/A130374/A130375 - DrNetworkSegmentationRemote KB 13513 - Alert - A160157 - FileServerVMcoreDetected KB 13889 - CRUD operations are not permitted on a stale Volume Group KB 13956 - CMSP enabled PC upgrade fails with error "Max retries done: unable to push docker image msp-registry.cmsppc.nutanix.com:5001/alpine:3.12" KB 13983 - Allow Out-of-band (OOB) snapshot for 3rd patry backups on RF1 VM KB 14177 - ESXi upgrade stuck at 14% due to vMotion taking longer than 15 minutes KB 14186 - Nutanix Files - Manual DNS verification fails even if the DNS entries are properly created KB 14195 - Acropolis service may crash with "Unexpected CAS error while trying to delete an
Below are new knowledge base articles published on the week of January 15-21, 2023.KB 10596 - Prism Central Pre-Upgrade Check: test_prism_central_cmsp_enablement_check KB 14074 - LCM Pre-check: test_empty_ilo_task_queue KB 14105 - After AOS upgrade to 6.x, users with the Prism Element Backup Admin role are unable to use AD credentials to log in to Prism Element KB 14167 - AHV upgrade is blocked in LCM 2.5.0.4 or newer if running vGPUs VMs are found KB 14185 - Power state of VMs managed by the Nutanix AHV Plugin for Citrix on AHV clusters changed to "Unknown" after upgrading AOS cluster to version AOS 6.6 or newer KB 14194 - Metrics-enabled (vhostmd) VMs fail to start on AHV 20220304.x KB 14198 - UpdateVmDbState task can often be seen on AHV clustersNote: You may need to log in to the Support Portal to view some of these articles.
Below are new knowledge base articles published on the week of January 8-14, 2023.KB 12052 - NCC Health Check: Prism Central disaster recovery PE-PC version compatibility check KB 13210 - NCC Health Check: pcvm_reboot_check KB 13272 - Pre-Upgrade Check: test_max_pe_nodes_registered_to_pc KB 13301 - Pre-Upgrade Check: test_prism_central_run_cmsp_preupgrade_check KB 13666 - Pre-Upgrade Check: test_prism_central_pcdr_compatibility_check KB 13676 - NCC Health Check: category_threshold_alert KB 13766 - Prism Central Pre-check: test_event_audit_counts_under_threshold KB 13872 - Pre-upgrade check: test_prism_central_pcdrv1_enabled_check KB 13979 - Custer creation fails on HPE DX G10 Plus nodes if there are no disks installed in a storage controller KB 14076 - In microservice infrastructure enabled Prism Centrals, /home directory space can be exhausted due to docker folders KB 14130 - LCM is running in Offline mode KB 14142 - Nutanix Kubernetes Engine - Pod to service communication does not wo
I can’t create a new scenario it says: “Install country missing”There is no option for me to input the Install CountryOnly options i have are Scenario Name and Scenario Objects.
As title, Does AOS 6.x support Fault-Tolerance on ESXi 7.x ?If positive, does it support more than 4 FVTM in an ESXi 7 host ? I had read https://docs.vmware.com/en/VMware-vSphere/7.0/com.vmware.vsphere.avail.doc/GUID-57929CF0-DA9B-407A-BF2E-E7B72708D825.html . Can I change parameters to support more than 4 FTVMs in an ESXi host?
Hi, One of my employees deleted 2 VMs with the name like xxxxxxxxxxxB1234 and xxxxxxxxxxxxxB1235 and at the same time xxxxxxxxxxxxxT1234 and xxxxxxxxxxxxxxxT1235 were deleted. (Logs show that he deleted them like less than a minute later) He claims that he did not touch those two. Is it possible that two VMs can be corelated somehow and deleting one triggers deletion of another one? One was bench and one was test. The only difference in name was one letter.
Hi all, I’m wondering if someone can help me out with the first install of a nutanix CE single cluster with esxiThe installation went fine but, and followed this link to make it work. https://vmik.net/2021/01/26/nutanix-ce-install-esxi-2021/I see that only have 1 disk in the storagepool. Is there a way to add a disk into the CVM? or do I have to mount it first. can someone help me out? many thanks.
My hardware - NX-6035-G4 running 5.20.3.5 LTS and is no longer under support. It is one of 8 in the cluster. We use this as a backup target for our two other production clusters using protection domainsI was attempting to replace a satadom because I got this error satadom has worn out - PE cycles above 4500 or PE cycles above 3000 and daily PE cycles above 15.I followed the instructions on this page to get things started, but after 8 hours of trying to clone the satadom disk the process timed out - put the node in maintenance mode and detached it from the Metadata Store. I took the node out of maintenance mode to add it back to the Metadata Store, but the option has not popped up in PE. Here is the output from an NCLI host list. Id : Uuid : Name : IPMI Address : Controller VM Address : Controller VM NAT Address : Controller VM NAT PORT : Hypervisor Address : Host S
Hello Ma’am and Sir,Does anybody tried installing or imaging a single node using phoenix? please i need some help/procedure for this as I can’t access the links given in this page it only shows “Not Found or Fordbidden”: Appendix: Imaging A Node (Phoenix) | Nutanix Community…...specifically in “Installing the Controller VM and Hypervisor by using Phoenix” and “Imaging Bare Metal Nodes”…..I created bootable phoenix using USB then i am now in the root@phoenix but I bet it requires some commands to push but have no idea on the commands and theres no search result in the internet regarding that
Hi. I’m trying to use Nutanix collector 4.0 to collect Hyper-V sizing information, however the following error message is displayed during the scanning phase:“Bad http response returned from the server. Code: 500. Content”I tried to use it with Remote option and locally on a cluster member and the same happens.Running the precheck script it shows these failures on the picture: The user logged in is a Domain Admin and local admin, however the script returns a failure about the user rights. Does anyone know how can I enable the “Basic auth” stated on the script output?I also tried Collector version 3.5.1 with no success.Any tips?
tl;dr how to perform pre-upgrade checks from the command line?Hello Guys,I have a question related to the performing manual / cli upgrade (command line). Here’s the scenario:AOS with ESXi as a hypervisor, multiple ROBO offices / Dark Site (no access to the internet), Dark Site server with LCM, very limited broadband: 16Mbit with static shaping (4Mbit) for management purposes and 12Mbit for businessIt is very difficult to upload via AOS GUI 4-5GB of upgrade and .json file. 3-4% / 10minutes and upload task informing about a ‘server issue’. I’m 99% sure that it related to the broadband shaping.Solution is simple:upload tar.gz install/upgrade bundle, ‘untar’, install ( ./install/bin/cluster -i ./install upgrade )My question is: is any magic command to start pre-upgrade tasks with is normally presented in GUI if upgrade package is available? It will be very helpfull to start pre-upgrade check in-advance the main upgrade.I tried to find the command in lab:-- uploaded the bundle in ‘normal
Hi Our setup will look very similar to the attached diagram once it’s setup (taken from here - https://portal.nutanix.com/page/documents/details?targetId=Nutanix-Security-Guide-v6_1:wc-network-segmentation-intro-wc-c.html). We won’t have the separate DVS for the VM’s though as these will be on the same DVS with everything else 😊 What I'm trying to understand is what the vmk2 kernel interface is used for and why this must be on its own port group and can’t use the same port used by the CVM. I’ve illustrated what I thought should be possible with the blue line. Within Prism where you change the setting for the Backplane network, you must select a port group for the CVM and for the host. However, when you try to select the same port group for both host and CVM it gives an error and won’t let you do it. Does anyone know why they need to be on separate port groups even though they will be on the same Backplane VLAN? I hope that made sense
nutanix ERROR ipv4config.py:1811 Unable to get the KVM device configuration, ret 1, stdout , stderr br0-backplane: error fetching interface information: Device not found\n Anyone know what to do with this error. I turned on remote syslogging to see why a host has been crashing and I see this as one of the errors in the logs. The host is a Dell PE R630.Not sure if KVM means the host OS, since AHV basically is KVM or an actually connecting to iDRAC for Keyboard Mouse Video.
What metrics are considered when Acropolis Dynamic Scheduling makes a migration decision on AHV related to CPU usage? Does ADS consider only CPU usage on the Host? Or does ADS take into consideration other VM or Host metrics such as CPU Ready, Co Stop, and/or Steal Time?
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.