Get guidance, share wins, and ensure smooth Nutanix deployments.
Recently active
If you are replicating data through DR (Data Replication page in Prism), then you have setup schedules for snapshots so they can be copied to the remote site on scheduled time. When you select a “snapshot” filed in one of the protection domains you created in the Prism UI, one of the fields is “Reclaimable space”. You may observe an spinning wheel continuously and the word “processing” on this filed for some or all snapshots, but you may also notice that snapshot(s) has already been taken and done. So why the spinning wheel for this filed? This field is lazy-calculated by Curator during full scans and populated afterward, so it takes sometime (may be few hours) to show up. Until Curator finishes calculating the value, the field shows Processing in Prism.
Hi, We want to throttle replication bandwidth. In the Remote Site config in Data Protection, the option is to throttle B/W in MBps. Is this Mbps (megabits per second) or MB/s (Megabytes per second)? The unit implies megabytes per second (i.e. the "B" is capitalised), but that is normally a unit of throughput not bandwidth... Obviously there is an order of magnitude in difference so its important we interpret this correctly!
Hello, Lately I have been adding the Nutanix SCOM Management Pack version 2.4.0.0. From the start on it worked fine as I handed over the Cluster information via Nutanix Cluster Discovery. But now, a few weeks later, SCOM will not display any performance data of the clusters anymore. I am not able to find out where exactely caused this problem. I have different Clusters AHV and ESXI as well as multiple OSVersions running. Currently SCOM displays only one Cluster with it's information. From the other Clusters, there is no performance data shown. All other clusters were discovered correctely (and the same way) but won't show their data anymore. My situation looks like this - this dashboard is showing the clusters information. All other clusters (and their nodes) are not showing data - the dashboard stays empty. What could have happend that the data is not reaching SCOM or SCOM not displaying it anymore? Could it be too much Clusters on my system? I have 13 Clusters with (together)
Hi I have a concern with the data resilience in Nutanix Cluster about rebuild the data in 2 scenarios. When a node is broken or failure, then the data will be rebuilt at the first time, the node will be detached from the ring, and I can see some task about removing the node/disk from the cluster. The whole process will used about serveral minutes or half hour. It will last no long time to restore the data resilience of the cluster. When I want to remove a node from the cluster, the data will also be rebuilt to other nodes in the cluster. but the time will be last serveral hours or 1 day to restore the data resililence. Seems remove node will also rebuild some other data like curator,cassandra and so on. but Does it will last so long time, hom many data will be move additionaly ? and What the difference for the user data resilience for the cluster?
Nutanix AOS offers simplicity in managing traditional complex infrastructure tasks. From Virtual machine management, Storage operations, replication - and of course Cluster software and hardware upgrades. As Infrastructure admins, we are well aware of the operational pain points, when it comes to upgrading: Hypervisor Upgrades Storage OS upgrades Firmware Upgrades Management software upgrades the list goes on… With Nutanix One-Click upgrades, customers can upgrade software components and hardware components easily. Software and Firmware needs to be downloaded from Nutanix repositories - which is why it is important to understand what Network Ports are required to be open or can be opened on demand to check for upgrades. Following KB from Nutanix Portal lists the required network ports for different services and upgrade repos endpoints: Recommendation on Firewall Ports Config
I am looking for a something that i can setup in an automated task on a server to poll for any active replications for Protection Domains and if true to pull information and email it. The output i am looking for in the email would be something like below. Protection Domain : ProtectionDomainName Replication Operation : Sending Start Time : 03/11/2019 12:00:02 EDT Remote Site : RemoteSiteName Snapshot Id : 2918635 Bytes Completed : 444.08 MiB (465,653,447 bytes) Snapshot Size : 2.57 GiB (2,760,598,528 bytes) Complete Percent : 95.38689 If anyone already has something like this setup that would be awesome, my scripting skills are slim to none so any help would be awesome.
Hi all, I’m very new to Nutanix, and pretty new to Ansible. I’ve been tasked with updating / installing guest tools on any machines that need them, and they’d prefer to do it via Ansible. I’d like to be able to have Ansible use the uri module to grab the UUID of a given VM, or grab a list of UUID’s and the associated VM; however I’m having a lot of trouble parsing this information out in a way that Ansible can actually use it. Does anyone have experience with this? Or at least can tell me that there’s a better way to be doing this? Thanks!
I'm running Prism Central 5.7.1.1 and have recently upgraded 6 of our clusters to 5.5.7.1. Of those 6 clusters, 3 of them are still showing an Upgrade Status of 'Upgrading' in Prism Central. It's been a over a month for one cluster and I've restarted Prism Central appliance to no avail - has anyone else seen this?
I've seen servers that have active and inactive in replication . when I check the replication settings. there are inactive and active vm's in job. so i didnt understand that what is mean inactive vm's. can i delete it . if there is no replication in the other region, they take up storage space and if I delete them I can save space from storage space Also i have another question is on vcenter 6.5. i see free size is 4 tb but on nutanix mangement screen free size is 11 tb. why do i see it so different. If the virtual machine is deleted, there is a space recovered on the storage side.
Hey everyone, I am required to upgrade from ESX 6.7 U1 to U2 to resolve a bug that prevents me from going back more than 8-10mins to look at VM performance, however, Nutanix has only just now tested U3. My question is: Has anyone migrated to U3 yet and if so how has it been so far?
Need help to change the IP address , Subnet mask and gateway of CVM, IPMI, ESXi Hosts and vCenter AOS 5.0.4 ESXi 5.5.0 Build number - 1331820 ncc 3.1.2 vcenter 5.5.0 Build number - 2442329 total number of nodes : 5 1) what is the Order to change the IP address, subnet mask and gateway of 5 CVM, 5 ESXi Hosts, 5 - IPMI, 1 - vCenter 2) do i need to stop the cluster for CVM - IP, subnet mask and gateway change? 3) do i need to stop the cluster for ESxi Hosts - IP, subnet mask and gateway change? 4) is there any workaround to update the CVM - IP address, subnet mask and gateway without stopping the cluster 5) reason for stopping the cluster while changing the IP address, subnet mask, gateway 6) does this AOS version will support CVM and ESXi IP change 7) what has to be check before and after IP address, Subnet mask, gateway change of CVM, ESXi, IPMI and vCenter 8) What is the Procedure of changing the CVM IP address, Subnet mask and gateway (AOS 5.0.4/Esxi 5.5.0) 9) What is the Procedur
@Mutahir has already shared some insights on NCC checks in Keeping the Lights Green - NCC - Hardware Checks. Today I would like to bring up two important aspects of the tool. There may be a time where you receive an alert triggered by a regularly executed NCC check. Oftentimes the alert will have a reference to a KB article. You read the KB and it does not make any sense. Naturally, you raise a case with the Nutanix support team or commence the journey across vast space of the Internet in the search for an answer. The very first thing Nutanix support engineer will do is verify if the environment is running the latest version of NCC checks, and if it’s not, they will proceed with the NCC upgrade. More often then not, the alert will clear after the NCC upgrade. Why is it so? NCC is a powerful tool that is developed and maintained by a team of professionals. With their help the tool evolves and grows, more checks are introduced, issues are resolved and algorithms are improved. Thus it i
We are using a Prism Central Installation with several Clusters in different Sites. For the Administrators on Site, we want to configure a user on the site based prism elements instance, which is a viewing User, with the following additional permissions on the local cluster / Prism Elements: Viewing, Power on and off for all virtual Machines and also a one click shutdown of the whole cluster in case of power outages. We don't want to use the Prism Central User Configuration for following reasons: Prism Central is not accessible in case of line outages or power failures in our sites. And for security reasons and limited skills of the staff on site, I really appreciate not to give admin permission to the site staff. Any Idea, how to deal with this situation ? Thanks Oliver
Hi Team, We have multiple Nutanix clusters with Dell XC servers installed with ESXI servers. We found an user account "PTAdmin" present in Dell iDRAC servers. Is there anyway to disable the account from CVM or Hypervisor since we have 100+ nodes in the infra. Thanks.
Is there a minimum or recommended switch port buffer size? I plan on using eth0 and eth 1 to be connected to two different 10G switches but was wondering if there is a minimum buffer size or switch port requirement?
Below are the top knowledge base articles for the month of November 2019. KB 4141 - Alert - A1046 - PowerSupplyDown KB 4116 - Alert - A1187, A1188 - ECCErrorsLast1Day, ECCErrorsLast10Days KB 1540 - What to do when /home partition or /home/nutanix directory is full KB 7503 - G6, G7 platforms with BIOS 41.002 -DIMM Error handling and replacement policy KB 4409 - LCM: (LifeCycle Manager) Troubleshooting Guide KB 1113 - HDD/SSD Troubleshooting KB 4541 - Alert - A101055 - MetadataDiskMountedCheck KB 4158 - Alert - A1104 - PhysicalDiskBad KB 2090 - AHV | Host and Guest Networking KB 4519 - NCC Health Check: check_ntp KB 1888 - NCC Health Check: storage_container_mount_check KB 4188 - Alert - A1050, A1008 - IPMIError KB 1507 - Alert IPMI IP address on Controller VM was updated to ... without following the Nutanix IP Reconfiguration procedure, can be misleading KB 4273 - NCC Health Check: aged_third_party_backup_snapshot_check KB 3523 - How to create a Phoenix ISO or AHV ISO from a CVM or Foun
Below are new knowledge base articles published on the week of November 24-30, 2019. KB 8302 - Pre-Upgrade Check : test_is_hyperv_nos_upgrade_supported KB 8303 - Pre-Upgrade Check : test_if_cau_update_is_running KB 8499 - Security - Nutanix definitions for most common STIGs KB 8555 - Launching a blueprint by using the simple_launch API fails after Prism Central is upgraded to 5.11 KB 8616 - "Restore" screen under ASYNC DR is misaligned if entity have long name KB 8618 - PD: Trying to to 'deactivate-and-destroy-vms' operation got error 'Error: Unexpected application error kInvalidAction raised' KB 8619 - Genesis may not start with error 'Received multiple ips for interface bound to ExternalSwitch' KB 8621 - Alert - A400101 - NucalmServiceDown KB 8622 - Alert - A400102 - EpsilonServiceDown KB 8629 - Calm - Jenkins deployment is stuck at "Installing: ssh-credentials" and fails without error messages KB 8639 - AHV | Never-schedulable node CVMs are not shown in the VM in Prism. KB 8641 - De
Hi all, I have some questions that I’m trying to answer but … ;) So if you can explain to me or point me to a part of some resources # Questions Is it recommended or mandatory to configure containers as ReplicationFactor-3 when the cluster is RedundancyFactor-3 In case of ReplicationFactor-3, when reading, how many checks are done to validate data correctness? In a RedundancyFactor-2 only 1 failure is tolerated, the cluster will still work with (e.g) 2 Zookeeper. In RedundancyFactor-3 there is 5 Zookeeper, so why we can’t tolerate up to 3 failure? What are the limitations for which it is not possible to migrate VMs between containers without the export/import method? How the cluster will behave in case of network separation issue (e.g. 4 nodes can communicate and 4 other too)? If I a have 2 Guest VM in the same Vlan, will they communicate through the OVS br0 or the traffic will go till the external switch and come back to the cluster? With the bond0 (br0.up) interface having 2 links
We have had a couple of instances recently when making network changes that have affected our clusters. This caused a restart on the lead host due to it detecting a network loss and then resulted in system outages. The cluster is configured with dual networks ports in active and passive mode and the understanding was that it would switch if any change or failure was detecetd without producing error events and systems down.
Hi all, A network security audit on a customer infrastructure reported a vulnerability on the cerebro http (port 2020) who is open on http in every CVM and without any security prompt.Some sensitives informations are visible : - AOS version : el7.3-release-euphrates-5.10.7-stable-... - VM Names - Protection Domain names - Witness ip address - ... Is there’s a way to secure this component ?
Hi, As per title i am wondering about encryption when we are using ESXi hosts and Nutanix together. We would like to encrypt the VM's and maybe Data-At-Rest encryption. We have currently 3 nodes in a block with 3 ESXi hosts and i wonder what the best practice is about encryption? Is it possible to encrypt the data both via Nutanix and vmwar software/KMS? Is it possible to encrypt only the VM's via VMware encpryption method and use the KMS directly via VMware or should i use the prism management? Does it matter in a Nutanix way if you choose to set up the KMS and encrypt the Data from the vSphere instead of doing it in the Prism mangement? Thank you in advance.
Hi In my previous publication, publish version v1.2. (https://next.nutanix.com/scripts-32/nutanix-tools-for-ahv-v1-2-32075) It was updated to v1.3 and it already extracts new information. The status of the NGT TOOLS, description, etc. and many new key validators to avoid problems. Additional I add a small script to obtain the information of the connection of the cluster to the TOR switch and the physical ports. Please check the github for more information. https://github.com/dlira2/Nutanix-tools-for-AHV I use the script in large accounts with more than 1000VMs and it works optimally.
Hi, I got a question during the test. Window VMs(C: drive) , linux VMs(df -h) capacity different those capacity in the prism(VM-table). Not all VMs are like this, only a few are like this. These VMs erased a lot of data and got a lot of capacity. The Curator has since been executed, but the capacity has not decreased inf prism-vm-table. Why can't Prism return a VM's erased data capacity value? How do I get Prism to read this capacity? Window -> Server 2016 Linux -> Centos 7.X Thank you.
Hi This is my first blog entry and therefore I would like to introduce myself briefly. My name is Omero Muscio, I have been working at Amanox Solutions in Switzerland as Presales Engineer for a little more than half a year. This means that I mainly take care of the customers' technical concerns, create sizings, carry out PoCs and go to the customer to carry out post-sales activities. How-To Cancel Stuck Move Migration Plan: With Nutanix Move 3.3.1, we had the error twice at our customer's site that the migration plan did not fail after manual abort, but hung up. With the help of Nutanix Support, we were able to quickly find a solution to this problem. Log in to the AHV VM via SSH or in the console. Username: admin PW: nutanix/4u enter rs,this is similar to sudo or su From now on you are logged in as root@move. Now we connect to the Postgre database: postgres-shell to open the PG-Shell psql -d datamover to get into the corresponding PSQL shell With the help of the command select mpuui
There is an ever-increasing chance that the next time you will need to add a node to the cluster or build a brand new cluster using new nodes – they will be bare metal nodes. What that means is that there is neither AHV nor CVM pre-installed on the node (unlike with factory imaged nodes). The process of imaging differs slightly from handling a factory imaged nodes hence this post. Bare metal nodes might come with a Discovery OS - small footprint software which allows for the node to be discovered by AOS 5.11 and later and thus imaged by an existing cluster CVM. When building a new cluster standalone (bare metal) imaging is performed from a workstation with access to the IPMI interfaces of the nodes in the cluster. Imaging a cluster in the field requires first installing certain tools on the workstation and then setting the environment to run those tools. This includes setting up Foundation VM and uploading Nutanix software images to the Foundation VM. Now follow the guide and let the i
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.