Get guidance, share wins, and ensure smooth Nutanix deployments.
Recently active
Hi,First time attempting to install Nutanix. I’m experienced with VMWare, but I wanted to try out Nutanix, mainly as I read it will allow me to create a single-node cluster, using locally attached NVMe disks without the overhead (headache) of vSAN (that doesn’t work as a single-node).Running through the installer, I see I have to specify an AHV (equivalent of ESX?) boot disk and CVM boot disk that requires a disk of at least 200GB. If I understand correctly, CVM is the equivalent of vCentre Server (VCSA).Firstly, can I install this CVM somewhere else?Secondly, I don’t have a separate disk of >=200GB that I want to waste on CVM. I have a 128GB SSD that I would have used for data, but the installer won’t let me and I can’t use it for CVM either because it’s too small. If I can’t farm off the CVM to another device, is there a way to override this limitation? Even fat old VCSA only uses 72GB across its 13 disks. What is in CVM! These are my disks. As you can see, to install CVM on its o
Does anyone knows size of VM’s snapshot file? I am concern about disk space, but I also need to back up. Is there formula for calculate snapshot file size approximately?
Hi! I've got several questions about the usage of v2 API which I have not been able to solve by myself after looking a little while.on some endpoints, for example, GET /alerts, GET /vms and a few more things are functioning for me as expected, but on other endpoints, such as POST /vms/{uuid}/set_power_state I am receiving the following response:{ "message": "Access is denied", "detailed_message": null, "error_code": { "code": 1100, "help_url": "http://my.nutanix.com" }}any idea what's causing it? I've been searching for detailed information on your error code (such as 1100 listed above) and couldn’t find. Is there any documentation to the error codes that you can send me a link to?Also, in a few of the API endpoints I couldn’t understand exactly what is the expected query parameter to be sent.2 examples of endpoints I felt I am lacking knowledge what exactly meant to be sentGET /vms filter: Filter criteria - semicolon for AND, comma for OR. where can I find what counts as fi
Hello. I`m installing AOS in FoundationVM. I failed while installing AOS.The failure log is as follows: *************************************Mount Information********************************Token 19936: drive path /home/nutanix/foundation/tmp/sessions/20210105-165854-9/phoenix_node_isos/foundation.node_192.168.1.124.iso mounted to SP 192.168.1.91 by user USERIDUmount successful.stderr:2021-01-06 01:47:08,796Z ERROR Exception in <ImagingStepInitIPMI(<NodeConfig(192.168.1.124) @1b50>) @0210>Traceback (most recent call last): File "foundation/decorators.py", line 77, in wrap_method File "foundation/imaging_step_init_ipmi.py", line 269, in run File "foundation/imaging_step_init_ipmi.py", line 180, in boot_phoenixStandardError: Failed to connect to Phoenix at 192.168.1.1242021-01-06 01:47:08,805Z ERROR Exception in running <ImagingStepInitIPMI(<NodeConfig(192.168.1.124) @1b50>) @0210>Traceback (most recent call last): File "foundation/imaging_step.py", line 161,
Below are new knowledge base articles published on the week of January 3-9, 2021.KB 10280 - A130149 - Guest Power Operation through NGT Failed. KB 10316 - Degraded performance on Lenovo hosts running AHV hypervisor as CPU runs at low frequency speed due to BIOS P-State configured mode failing to interact correctly with the hypervisor driver KB 10502 - Prism support for SNI (Server Name Indication) KB 10521 - Acropolis service stalling on all nodes in AOS 5.17.x. UEFI VM stuck in power-on state KB 10542 - Protection Domain schedules following DST time changes is not affecting hourly schedules KB 10550 - Leap | Recovery plan failover of VM fails with message "NGT reconfiguration failed. error detail: INTERNAL_ERROR: ErrorCode: 9" in the UI however VM has migrated and powered offNote: You may need to log in to the Support Portal to view some of these articles.
02/09/2021 Update: 3.10.1 has been released and contains the fix for this problemAn attempt to add a cluster with NVMe drives in the X-Ray 3.10.0 fails immediately.While checking the log on the X-Ray VM the following can be observed in the xray.log:ERROR Error occurred while performing final discovery - Invalid value for `type` (SSD-PCIe), must be one of ['HDD', 'SSD', 'Unknown'] That issue is specific to the version 3.10.0 of X-Ray and affects only the clusters with NVMe drives. It is a software bug that is going to be fixed in the upcoming release 3.10.1.The workaround is currently to use the earlier versions. X-Ray 3.8 is validated to work correctly.Earlier versions of X-Ray can be downloaded from the following page: https://portal.nutanix.com/page/downloads?product=xraySimply click on “Other versions” to see the previous versions.
About Xi Frame:Xi Frame is a secure cloud platform that lets enterprises and independent software vendors (ISVs) deliver applications, desktops, and software-defined workspaces to users. Users only require a connected device with a modern web browser. There are no clients, downloads, or plugins to install.Please follow this KB article for a step-by-step procedure for setting up and configuring your Cloud Connector Appliance manually.Note: CCA and WCCA VMs are automatically created in PC versions above 5.11 when deployed. Requirements Nutanix cluster running AHV with Acropolis Operating System (AOS) with Prism Central 5.10 or newer Frame Agent and Frame Cloud Connector Appliance (which you can find here) A Xi Frame subscription (sign up through https://my.nutanix.com/) Review the Network Configuration Requirements documentation which outlines required network protocol and port configurations. Note:If the CCA is already deployed, then just follow the steps in the Frame Documentatio
Switch engineer changed vlan (vlan id). Is it possibly auto update in nutanix arp table? if not, what kind of command require to do?
Flash Mode is a great feature ensuring that VM workloads remain within the flash (SSD) tier of storage. Once flash mode is enabled for a virtual machine, all of the disks associated with that VM (including any future created disks) automatically get added to the flash tier.However, sometimes having so many disks within the flash tier can cause performance degradation for other VMs that are not configured for flash mode (but could benefit from using the flash tier of storage) or can cause the available flash tier space to be consumed too quickly. Further, it is sometimes not desirable to have all of the disks associated with a virtual machine contained within the flash tier.Accordingly and, though not available as a Prism Web user-interface (UI) modifiable option, individual VM disks can be configured to not use the flash tier even while the VM itself is configured for Flash Mode. The procedure for removing individual VM disks from the flash tier involves using the Acropolis Command-Lin
When I try to run the Nutanix foundation applet, Java tells me that the file does not meet high or very high security requirements and that it is missing a certificate. I can’t add it to the site exception list because the file: protocol is not secure. Some guidance would be helpful.
Network visualization Is one of those underrated but extremely useful features. Network visualization is a consolidated graphical representation of the network formed by the VMs and hosts in a Nutanix cluster and first-hop switches. Nutanix network visualization allows you to group VMs by power state or by parent hosts, group hosts by parent clusters or select a particular node from the cluster. Choose your layout based on the current needs. In Nutanix network visualization, every entity is interactive. For example, when you click on a VM you can see detailed information about its vNICs such as MAC address, IP address, VLAN ID, live statistics. Similarly, switch port will show information about MAC address, MTU size, live statistics. To enable Network visualization a set of requirements must be met: Configure SNMP v3 or SNMP v2c on TOR switches Enable LLDP or CDP on the first-hop switches. Network connectivity over SNMP port between CVMs and switch management IP addres
I am facing an issue with the sizer while generating a budgetary quote. Help please!
Can anyone explain the function of the command “nodetool -h 0 ring” and also the explanation of its output, because I often see nutanix support using this command, i just curious function this command :D
I am trying to execute a remote PS script in a windows server using the playbooks. Have kept the script very simple for now to test the execution, but I get the error “ Failed to execute action with error: Internal Error, Connection reset by remote peer.” every time its executed.Execution Start Time: 01/06/21, 5:21:32 PM Execution End Time: 01/06/21, 5:21:32 PM Result: FAILEDInputUsername: serveradminPath to Script: E:\NutanixPS\Nutanix-CVM-Post-reboot-checks.ps1Password: ********IP Address/Hostname: 10.32.135.100HTTPS: True Has anyone faced a similar issue and what is the fix?
Hi I deploy a 3 nodes cluster with ESXi 7.0 as the hypervisor successfully on AOS 5.15.3.But when I run the diagnostic.py to check the cluster performance. It always failed to boot the diagnostic VM. I tried 3 times and all is same. from the log, seems the CVM cannot reache the diagnostic VM when check the boot status.I login to vCenter and check that there is no IP address for this VM in the summary page. And I can login the diagnostic VM from the VM console, I can check the ip address which is 192.168.5.253, but I cannot ping 192.168.5.254 and 192.168.5.1. also cannot ping 192.168.5.253 from CVM.From this information, I think the CVM cannot reache the diagnostic VM and timeout, so the diagnostic.py will fail after timeout. Is there any idear for this test? I can run thie script successfully in ESXi 6.7.
Hello. i Have a 4 nodes Lenovo HX3320. When i want to install AOS it fails. Foundation 4.6AOS 5.15.4 I updated bmc firmware and reset it. But it didnt work.Can you help me with this problemhere is the log 2020-12-10 12:38:39,191Z DEBUG Setting state of <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0> from PENDING to RUNNING2020-12-10 12:38:39,197Z INFO Running <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0>2020-12-10 12:41:01,398Z DEBUG Cache HIT: key(<function common_validations at 0x03D71230>_()_{'global_config': <foundation.config_manager.GlobalConfig object at 0x054D0E90>})2020-12-10 12:41:01,405Z DEBUG Setting state of <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0> from RUNNING to FINISHED2020-12-10 12:41:01,410Z INFO Completed <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0>2020-12-10 12:41:01,415Z DEBUG Setting state of <GetNosVersion(<NodeConfig(192.1
I have 20 clusters under management of our Prism Central and would like to know the versions of NCC installed in each. Is there a way to singularly collect this info across all PE clusters? Perhaps in nCLI even?
So guys I’ve got a flexlm license server running Windows Server 2019 that I am trying to use Move to migrate into my AHV cluster. I tried migrating it once and it blew up the license server since I guess it builds a unique machine ID using a disk and perhaps CPU ID. Any idea on how to clone this in Move so when it comes across FlexLM and other license daemons are none the wiser?
This article contains IPMI commands for checking and setting interfaces to dedicated or shared mode. For example, after a BMC upgrade, the IPMI might not be accessible. So, you need to verify and change the interfaces to dedicated or shared mode. Note: To run ipmitool commands on an ESXI host, prefix all commands with a forward slash (/). Note: To run ipmitool commands from a remote system such as the CVM (Controller VM), add the "-I lanplus", "-H <IPMI IP>", "-U <username>" and "-P <password>" parameters to the ipmitool command. For example: nutanix@cvm$ ipmitool -I lanplus –H x.x.x.x –U ADMIN –P <password> <command> Quanta Platform Use these commands for an NX-3400 (Quanta) platform. All commands are executed dynamically and a restart is not required. Check the status. [root@host]# ipmitool raw 0x0c 0x02 0x01 0xff 0 0 An output similar to the following is displayed.1100 :00 - Shared port 1101 :01 - D
In some cases, you might have to permanently remove a physical node / host from a Nutanix cluster. There are two scenarios in node removal. Permanently Removing an online node Removing an offline / not-responsive node in a 4-node cluster, at least 30% free space must be available to avoid filling any disk beyond 95%. You cannot remove nodes from a 3-node cluster because a minimum of three Zeus nodes are required. Some Points to consider before initiating node removal: Sufficient Disk space available on other nodes in the cluster User Virtual Machine relocation (if required) Any software upgrade should not be running Checklist on verifying cluster health status Data resiliency is “OK” (green) in Prism Run a complete “ncc report” either from prism or CVM cli: ncc health_checks run_all Depending on the size of data, node removal can be lengthy process, which involves relocating data from the node to other healthy nodes in the cluster. Node removal also remo
Hi,I am implementing a Nutanix Cluster with Lenovo HX5520, when I register the Prism in vCenter the process appears as completed but I cannot create virtual machines by Prism Element and when running the NCC I return communication errors with the vCenter is not stabilized. I did the communication test of the CVMs with the vCenter on ports 80 and 443 and the connection worked successfully. when checking by cli I see that the connection is ok, but it does not provide me with the settings of vcenter follows the difference from another implementation any idea what it might be?
Below are new knowledge base articles published on the week of December 27, 2020-January 2, 2021.KB 10019 - Alert - A110024 - AwsDefaultUVMSecurityGroupNotFound KB 10507 - Nutanix Move | VM migration fails if Hyper-V VM has Fibre Channel disk controller attached KB 10528 - Era - Operation failed with Internal Error after changing Cluster Account Password KB 10529 - [ Karbon ] How to configure email alerts for a Karbon kubernetes cluster KB 10530 - NCC-4.0.0 : Health Server logs might fail to rotate and fill up /home partitionNote: You may need to log in to the Support Portal to view some of these articles.
Below are the top knowledge base articles for the month of December 2020.KB 7503 - NX Hardware [Memory] – G6, G7 platforms - DIMM Error handling and replacement policy KB 10475 - LCM 2.4 inventory failure - [SSL: UNKNOWN_PROTOCOL] unknown protocol (_ssl.c:618). Not fetching available versions for module KB 4141 - Alert - A1046 - PowerSupplyDown KB 1540 - What to do when /home partition or /home/nutanix directory on a Controller VM is full KB 1113 - HDD/SSD Troubleshooting KB 4409 - LCM: (LifeCycle Manager) Troubleshooting Guide KB 4158 - Alert - A1104 - PhysicalDiskBad KB 2090 - AHV host networking KB 4519 - NCC Health Check: check_ntp KB 2473 - NCC Health Check: cvm_memory_usage_check KB 4116 - NX Hardware [Memory] – Alert - A1187, A1188 - ECCErrorsLast1Day, ECCErrorsLast10Days KB 4273 - NCC Health Check: aged_third_party_backup_snapshot_check and aged_entity_centric_third_party_backup_snapshot_check KB 6945 - How Upgrades Work at Nutanix KB 1863 - NCC Health Check: sufficient_disk_s
What is happening to my 2 node cluster during a failover or an upgrade?What does a recovery process look like after a node failure?If you are wondering the above, we have the answer for you! You can monitor the progress of your 2 node cluster in these situations through Prism Element.To monitor node recovery progress after failover: Registering a witness is highly recommended to help the cluster handle the failover situation automatically and gracefully. Stand-Alone mode: A failed node would trigger cluster to transition into stand-alone mode during which the following occurs: Failed node is detached from metadata ring. Auto rebuild is in progress. Surviving node continues to serve the data. Heartbeat: Surviving node continuously pings its peer. As soon as it gets a successful reply from its peer, clock starts to ensure that the pings are continuous for the next 15 minutes. If a ping fails after a successful ping, the timer will be reset. Prism Element Home page shows Critical
Prism Central includes machine-learning capabilities that analyze resource usage over time and provide tools to monitor resource consumption, identify abnormal behavior, and guide resource planning. These tools include VM "right sizing" where VMs are analyzed and those that exhibit inefficient profiles are identified. Anomaly detection to record when performance or resource usage is outside an expected range based on learned VM baseline behavior. "Smart" alerts that trigger when specified anomalies are recorded. Reports that summarize cluster efficiency. VM Right SizingIt is useful to look at the profile of your VMs when analyzing problems in a cluster or assessing future resource needs. This can help you identify VMs that are not optimally configured such as ones that consume too many resources, are constrained, are over provisioned, or are inactive.Anomaly Detection:The right sizing feature identifies inefficient VMs that fit one of the profiles described as below: Bully VM : A
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.