Get guidance, share wins, and ensure smooth Nutanix deployments.
Recently active
Good Morning:We are in the process of deploying new NX-8235-G7 nodes that are connected to a pair of Cisco Nexus switch. Everything seems to be working well, with the exception of the switch occasionally receiving RX Pause frames from the servers. While this may not necessarily be a problem, I am doing my due dilligence.The switch ports the nodes are connected to are configured as follows:Connected as active/standby NICs Connected via 1 meter 10GE DAC cables 802.1Q trunk Speed/duplex auto-negotiation Flow control auto-negotiationThe nodes themselves are configured with the default network configuration that comes from the factory with the exception of anything that is necessary to support the use of ESXi and our network environment.Any suggestions or validations of how the switch ports should be configured? Thanks!
Nutanix offers a distributed data and control plane, so it’s fairly easy to start, stop and graceful shutdown a cluster. Even, in abnormal / dirty shutdowns, Nutanix cluster has powerful self-healing capabilities - as all data & meta-data is distributed across the cluster which significantly reduces the chances for data corruption or data-loss. However, as a Nutanix Cluster hosts business critical data and applications, it is important to ensure all services stop in a graceful manner and all data + meta-data is consistent. This allows the cluster to be restarted in a healthy - usable state later.There can be several reasons to gracefully shutdown a running Nutanix AHV & AOS Cluster. When shutting down a Nutanix Cluster, following order needs to be followed:User VMs Shutdown AOS Shutdown - Data services / Cluster components CVM Shutdown Hypervisor ShutdownIn this post, we will focus on the shutdown process for a Nutanix Cluster running with AHV (Acropolis Hypervisor).Points to C
A piece of good news to begin with: it is possible to perform a non-disruptive memory upgrade on one node at a time.Start with a little bit of prep work. Check cluster status. Download and install the latest version of NCC (Nutanix Cluster Checker) and execute all health checks. Resolve any issue found. Identify the node you are performing the hardware maintenance on by turning on the chassis identifier lights on the front and back of the node that you will remove using one of the following options. The lights turn on for four minutes unless you change the value. Now you can migrate VMs of the node. Shut down CVM and the host. Replace or add DIMMs.For detailed outline of the process please refer to KB-1623 HW: Upgrading Physical Memory
Hi Team, Please let me know if you can share the Detailed Migration plan from ESXi to AHV. Please share the document or the link. Regards.
Hi All I updated ahv for the prism central and now it does not show the cluster licenses - how do I correct this? ThanksEric
hello guys, I am new in Nutanix I want to learn Nutanix , take exam and I need virtual environment for testing is there anyone can help me.
Let’s say you are managing an Infrastructure and have a range of networks defined in your environment and now you have to delete a few.You go to Prism, try deleting the network but get a generic error.So what’s going wrong here?Why can’t you delete a network?Well, one possibility is that if there are NIC connected to the network, it won’t let you delete the network, kind of like a guard-rail.So what can you do to mitigate it?Try giving the following KB a read KB-8234 Want to know more about AHV Networking? AHV Networking Best Practices Still confused about AHV Networking?Drop a comment and let’s start a discussion.
Hello,My client has this alert "Active Directory Domain Contoller(s) or DNS servers configured on the UVMs in the cluster" due the fact that he moved his domain controller on the nutanix cluster. His environment is 100% Hyper-V, and he is totally aware that SMB3 share of the nutanix cluster, requires authentication from the domain. In order to avoid that my client created ISCSI volumes and presented them to his Hyper-V environment, on which he moved his domain controller, trying to avoid that type of failure if everything goes down after a power failure and when everything comes up, the domain controller to be able to boot prior of the authentication.Please let us know what's the best approach for this matter and what's your recommendation for this kind of setup, especially when all the domain controllers are virtualized.Regards,Adrian
Every production infrastructure knows the importance of load balancing the network traffic to increase efficiency.Let’s say you have multiple links in your environment and want to use the potential of all the links or want to have a backup configuration in case a link fails, load balancing will come to your rescue.Today we will talk about two load-balancing modesActive-Backup Balance-slbTo know more about the load balancing configuration and AHV networking in detail, give the following document a read AHV Networking Best Practices Guide So how to make a decision regarding active-backup and balance slb?This comparison might help youAdvantages of Bond Mode for the active-backupDefault bond mode is active-backup. One interface in the bond carries traffic and all the other interfaces in the bond are used only when the active link fails. Active-backup is the simplest bond mode that easily allows connections to multiple upstream switches without any additional switch configuration. Disadva
LACP configuration is always a tricky business for every administrator and network engineer and enabling and verifying the configuration of LACP on NX Nodes should not send chills down your spine.First, let’s understand why we need LACP What are the advantages of LACP? A single user VM with multiple TCP streams could use up to 20 Gbps of bandwidth in an AHV node with two 10 GB adapters. A traffic-hashing algorithm such as balance-TCP can split traffic between multiple links in an active-active fashion. Because the uplinks appear as a single L2 link, the algorithm can balance the traffic among bond members without any regard for switch MAC address tables. With LACP, multiple links to separate physical switches appear as a single layer-2 link.Note: To use multiple upstream switches, you must configure MLAG or vPC on the physical switch Points to seriously consider before jumping towards configuring LACP Read the following a guide to understand AHV Networking better AHV Networ
Below are new knowledge base articles published on the week of November 3-9, 2019.KB 7563 - NCC Health Check: ofpfmfc_table_full_check KB 8279 - Alert - A130116 - Automatic Promote Metro Availability KB 8326 - AHV | How to convert VM running Windows 10/2016/2019 from BIOS to UEFI KB 8413 - Alert-A1087-Metadata Volume Snapshot Persistent Failure KB 8451 - Move Migration Plan fails due to duplicate VMUuid KB 8477 - NCC Health Check: remote_site_in_same_datacenter_check KB 8485 - Alert - A130088 - Failed To Snapshot Entities KB 8487 - Alert - A130197 - VmRestorationFailed KB 8492 - Nutanix Files - While trying to update permission using Windows Explorer "Enable inheritance" is not working KB 8495 - Missing vnics after AHV Protection Domain Failover. "Failed to find a network: Unknown network" error KB 8504 - Pre-upgrade check: test_vlan_detection KB 8506 - Era - Unable to register Cluster with due to non-UTC timezone on Era Server VM KB 8520 - Alert - A130199 - VM Migration Failed KB 8521
Suppose, it is discovered that a Nutanix cluster is configured with an incorrect time zone. While like all good things in life this requires a little planning, fear not – this can be helped!First things first, what to expect:Timestamps of Nutanix logs events will remain in the incorrect time zone until cluster services are restarted or the CMVs are rebooted. Cluster services restart on a cluster with active workload may impact availability of the workload. Changing the timezone for PE and PC is the same procedure, is not disruptive, but it requires a reboot of CVMs. Reboot can be performed on one CVM at a time. For a seamless experience it is recommended to evacuate VMs from the node prior to CVM reboot. Verify cluster Data Resiliency health prior to restart of cluster services or CVMs. Ensure Data Resiliency comes back healthy prior to proceeding with another CVM reboot. It is necessary to remove DR schedules prior to the change. Schedules can be re-created post change. Command for ch
AOS is in 5.10.8, We need to install NGT on Windows Server 2016 with only VSS feature enabled and not SSR. However, after installation, even though the SSR feature is not enabled, there is the SSR icon on the desktop. It can be opened and bring us to “http://localhost:5000”. I can understand that the feature is not active but I would like to avoid having it displayed on the desktop. How to avoid to have this on the desktop ? (I’m looking for another way than deleting the icon manually after NGT installation). What about this local web server ? Could it be stopped ?
When it comes to migrating your environment into AHV the recommended way would be to use Nutanix Move. Much being said and written and more is coming. Nutanix understand however that using Move may not be an option for whatever reason. Freedom of choice is important after all, right? In case you are looking for an alternative way to migrate your VMs’ disks into AHV then keep reading.The article referred to in this post covers three scenarios:Source virtual disk files can be accessed directly by the Acropolis cluster over NFS (or HTTP). This scenario is generally possible when: The source environment is a Nutanix ESXi cluster. The source environment is a Nutanix Hyper-V cluster. The virtual disk files are hosted on an NFS server and the Acropolis cluster can be provided access to the NFS server export. (Common in non-Nutanix NFS based ESXi environments.) Previously exported virtual disk files are hosted by any NFS or HTTP server and the Acropolis cluster has access to them. Source vi
Cluster data usage grows and occasionally it grows more rapidly then planned and expected before a new node to be added to handle this growth. Nutanix offers data Compression and Deduplication to hold more data in the container, by reducing the stored size and avoiding duplicate data, respectively. Usually Compression should be used first, as de-duplication is recommended only on some specific scenarios (please, check documentation links provided below). Compression: ============ You can enable compression on a storage container. Compression can save physical storage space and improve I/O bandwidth and memory usage which may have a positive impact on overall system performance. The following types of compression are available. - Post-process compression: Data is compressed after it is written. The delay time between write and compression is configurable, and Nutanix recommends a delay of 60 minutes. If compression is enabled in "Post-process", then existing data will also
I have had a vm fail and now I need access to its storage drive, is there a way to mount its storage drive to a new vm or browse and recover data?
Considering using Nutanix Move to migrate your environment to AHV but confused regarding the compatibility and firewall requirements Let’s break the information to help you plan your migration smoothly The following guide lists the firewall requirements and the ports to be opened for the Nutanix Move Move firewall Guide So what should be the AOS version and the supported Source Hypervisors which are supported and recommended by Nutanix? Go through the compatibility matrix below to give you a better idea regarding the compatibility of Nutanix Move with a different hypervisor Move Compatibility Matrix Are there any limitations for Migration? Yes there are few limitations which should be kept in mind before using Move Move limitation Have a query regarding Nutanix Move? Start a conversation in the Move Application Migration forum
In our previous post “Time Synchronisation on Nutanix Cluster”, we highlighted time sync importance and some commands to check status of time-sync on a Nutanix Cluster.In this post, we will briefly go through some important recommendations in selecting a time-source for your Nutanix Cluster.Nutanix recommends using at least 5 stable time sources that have a high degree of accuracy and that can be reached over a reliable network connection. Generally, the lower the stratum of an NTP source, the higher its accuracy. If lower stratum time sources (e.g. stratum 0, stratum 1) are difficult to gain access to, then higher stratum time sources (e.g. stratum 2, stratum 3) may be used.For more details, see Nutanix Recommendation for Time Synchronization. For a list of Stratum One Servers: http://support.ntp.org/bin/view/Servers/StratumOneTimeServersIf you have to select a Off-Site time-source, following is a good selection of points to consider:http://support.ntp.org/bin/view/Support/SelectingOf
AHV VM High Availability (HA) is a feature built to ensure VM availability in the event of a host or block outage. In the event of a host failure the VMs previously running on that host will be restarted on other healthy nodes throughout the cluster. The Acropolis Master is responsible for restarting the VM(s) on the healthy host(s). But we already know that, right? First, let us be reminded that there are three types of AHV High Availability configuration within AHV Cluster: Best effort Reserved Segments Reserved Host (only available via acli and not recommended in AOS 5.0 and newer) To read about each type as well as to find configuration steps, logs location and common issues with their explanation please read KB-4636 AHV | VM High Availability (HA). For a refresh on AHV VM HA: Acropolis Virtual Machine High Availability Resources Nutanix University: Tech TopX: VM High Availability in AHV Prism Web Console Guide - Virtual Machine Management - VM High Availability in Acropolis Tech
Hello, I'm trying to clone from a known UUID (our template) but I keep getting a 400 code for "Bad request" Ex : (Using Postman for this example, but have also tried with POSH and had the same result) POST https://1.1.1.1:9440/api/nutanix/v3/vms/{UUID}/clone Basic Auth - admin/pw Headers - Content-Type:application/json Once I send the API call, this returns : { "api_version": "3.1", "code": 400, "message_list": [ { "message": "Bad request.", "reason": "BAD_REQUEST" } ], "state": "ERROR" } Any advice / help would be appreciated. This has been boggling me as to why it'd throw this error. Thank you!
Hello Masters, Today I have this alert: System Non-Root Partitions Space Usage HighI read this KB https://next.nutanix.com/installation-configuration-23/cvm-non-root-partitions-space-usage-high-clear-space-safetly-33597and check my filesystem and all files inside of this directories are ok (small sizes):/home/nutanix/data/cores/ /home/nutanix/data/binary_logs/ /home/nutanix/data/ncc/installer/ /home/nutanix/data/log_collector/My problem look to be at the file “/home/nutanix/data/logs/ncc_log_collector.log”. It has actually this size (14 and 6 GB): $ allssh du -h ~/data/logs/ncc_log_collector.log================== 192.168.150.92 =================5.7G /home/nutanix/data/logs/ncc_log_collector.log================== 192.168.150.94 =================14G /home/nutanix/data/logs/ncc_log_collector.log================== 192.168.150.96 =================238M /home/nutanix/data/logs/ncc_log_collector.logthe file don’t show error, only INFO. It safe to delete it ?? can I delete it using “rm /home/n
Thinking of migrating your workload from a different hypervisor to AHV but confused regarding the timeline. Migration of large data is always a time-consuming process and should be planned effectively. Nutanix presents you with an estimated time taken for migration of different workloads with a different configuration. Planning your maintenance window for the migration? Go through the link mentioned below to get an estimate idea regarding the time taken for migration and plan accordingly. Move Migration Guide Have a query regarding Move? Start a conversation in the Move Application Migration forum
Hi everyone, I have an issue with a newly purchased NX-1175S-G6 that need to deploy at the EU environment.The system doesn’t allow to raise a case to the support portal.The installation just can’t seems to pass through the foundation stage.The latest place it stuck is at the screenshot. That is still just a small part. Before this when I try to put in VLAN to setup, it just can’t get through it. Anyone can guide me to a correct direction? Thanks
Occasionally you will need to deal with nutanix support who needs access to your cluster but not over a webex or zoom sessions. You can always setup a remote support channel so nutanix SREs can perform troubleshooting in the background, leaving you free to work on other more important issues in your environment. Simply run:To check if it is already set up:nutanix@cvm$ ncli cluster get-remote-support-statusTo start itnutanix@cvm$ ncli cluster start-remote-supportTo close itnutanix@cvm$ ncli cluster stop-remote-support For troubleshooting instructions, please refer to public nutanix KB titled: “Nutanix remote Support Channel Troubleshooting Guide”.
While at the time of posting this article neither the process referred to nor Windows 2003 OS itself are supported within AHV environment, Nutanix understands that there are situations where customers might find themselves in a position of not being able to move away from certain OS version.To help customers with the process of migration of Windows 2003 servers from ESXi to AHV we share this post by Artur Krzywdzinski where he explains the process in details. We would like to thank Artur for sharing the solution.Share your own ideas and processes that worked (and did not) with community - help someone, encourage cooperation!Please note that Nutanix Xtract referred to in the post is currently known as Nutanix Move.vmwaremine.com: Migrate Windows 2003 to Nutanix AHV by Artur Krzywdzinski
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.