Get guidance, share wins, and ensure smooth Nutanix deployments.
Recently active
Below are new knowledge base articles published on the week of May 30-June 5, 2021.KB 10959 - Unable to deploy Azure Cloud Connect from Prism UI KB 11329 - Linux VM performance troubleshooting scenario - ksofirqd process utilizes 100% of a vCPU in Linux VM due to a high number of iptables rules. KB 11368 - Objects Service Manager (aoss_service_manager) service on PC crashes due to expired certificate KB 11372 - LCM HBA upgrade to PH16.00.10 fails with "kLcmUpdateOperation for release.smc.gen11.hba.hba_LSISAS3008_2U4N_2U2N.Skylake.update on X.X.X.X with ret: -1, out" KB 11383 - Network Segmentation: "Interface eth2 is not present on cvm" Error Seen When Enabling Backplane Network KB 11398 - "Invalid network selected" message returned when editing network configuration of a Nutanix Files cluster KB 11404 - Expand Cluster may fail when running on an unqualified ESXi build KB 11438 - LCM upgrade for disk firmware fails if there is a USB device in the node KB 11444 - Enabling Cloud ConnectN
Is AOS single node cluster upgrade an disruptive action?Thanks!
Hi,I have three identical nodes (basic Lenovo servers) but one performs really slowly compared to the others. The disks aren’t the same across all three though. I wonder if anyone has any ideas which difference might be causing the performance issues: Firstly the HDDs are 4TB in the two ‘good’ nodes and 2TB in the ‘slower’ one. I wouldn’t expect that to cause performance issues, just overall capacity will be limited? To be fair the 2TB is a WD ‘green’ and the 4TBs are WD ‘red’ drives. Secondly the ‘slower’ host has a Samsung Pro 840 256GB SSD which Prism reports as often having a higher latency than the HDDs! See the image below Are they known for being poor in this kind of environment? I’m not really stressing them. Any good logs to see performance issues? I can see the HDD activity light is permanently on on the slower host. Cheers,Steve
Beside the “classical” Monitoring with snmp or agent based version now we are talking about Metrics and the Monitorng of LiveData with the OpenSource Project Prometheus. The Virtualization is made with the powerfull Tool of Grafana.The access to the Data will be made with so called “Node-Exporter” which could we found under GitHUB. One is present for Nutanix! My Testcase in the homelab is build with the following items:NutanixCluster <- Prometheus <- Grafana192.168.10.80 192.168.10.123 192.168.10.100Nutanix CE Ubuntu 18.04 LTS Ubuntu 18.04 LTSPrometheus 2.2.1 Grafana 7.0.4GO 1.10RequirementsCreate NEW User in Nutanix Prism Central with VIEWER Rights2. Installation of Prometheus on Ubuntu 18.04 LTS with running GO!Good Sourcse will be available here or here.3. Installation ofvon Grafana 7.x on a second Ubuntun 18.04 LTSGood Source are found here. Start of Connection to NutanixWe download the GO Binary for the Nutanix Exporter to the Prometheus VM in the folder of GO/BINWe will tes
I am getting this error when adding my AWS environment to Nutanix Move toolAWS Authorization Error. RequestCanceled: request context canceled caused by: context deadline My move tool version is 4.0.0
Hello, all. We’re running vCenter 7 with AOS 5.15.x, and I’m learning about how VMware has now decoupled the DRS/HA cluster availability from vCenter appliance and moved that into a three VM cluster (the vCLS VMs).In the interest of trying to update our graceful startup/shutdown documentation and code snippets/scripts, I’m trying to figure out how to handle these vCLS VMs. They reside on the Nutanix shared storage, so I obviously would like to shut them down before gracefully shutting down the Nutanix CVMs/ADSF cluster as well as ensure the CVMs are up and cluster is good before allowing them to power back on using that storage. Evidently, these vCLS VMs are very aggressive about powering back on or recreating themselves once deleted, so I’m a little unsure what to expect.So with regard to powering back on the ESX hosts, I assume when I take them back out of maintenance mode the CVMs will be powered back on (or maybe I have to do that manually?), and after waiting a few minutes, I woul
While it is not as a straightforward process as we would like for it to be there is an option to add a NIC to your Move VM. Login to Prism Element Add New NIC to the Nutanix-Move appliance and select the network Launch the console of Nutanix Move appliance. Switch to root user Use vi editor or any other editor of your choice to open the file /etc/network/interfaces Add the second interface eth1 configuration in the format below based on DHCP/Static IP addressing. Restart the networking service.Please Note: If you are using Move 3.0.3 or above you can skip Step-7 and Step-8. That'll be taken care automatically. There will be an existing script named "start-xtract" under "/opt/xtract/bin". Overwrite that script with the one provided in the KB (see link below). Change the permissions for the script. Stop iptables and restart move services. Verify the new eth1 interface configuration using "ifconfig eth1". See KB7399 - Procedure to add a second NIC interface on Move v3.0.2 for detailed ins
Hello all..is there a list with model number somewhere of the Nutanix supported NICs? I am looking for the cheapest 1Gb dual NICs Nutanix supported that I can get from Amazon or ebay or the cheapest 10Gb..I was reading that supermicro Nic is supported but I don’t see model number so I was wandering if the Supermicro AOC-SG-12 is supported since that card cost about $35..my understanding is the CE and commercial support the same NICs now. My use case is CE for homelab. thanks
Below are the top knowledge base articles for the month of May 2021.KB 7503 - NX Hardware [Memory] – G6, G7 platforms - DIMM Error handling and replacement policy KB 1540 - [AOS Only] What to do when /home partition or /home/nutanix directory on a Controller VM (CVM) is full KB 4141 - Alert - A1046 - PowerSupplyDown KB 1113 - HDD or SSD disk troubleshooting KB 4409 - LCM: (Life Cycle Manager) Troubleshooting Guide KB 4158 - Alert - A1104 - PhysicalDiskBad KB 6945 - How Upgrades Work at Nutanix KB 2090 - AHV host networking KB 2473 - NCC Health Check: cvm_memory_usage_check KB 4519 - NCC Health Check: check_ntp KB 1863 - NCC Health Check: sufficient_disk_space_check KB 3827 - Alert - A130087 - Node Degraded KB 7554 - Critical : Cluster Service: Aplos is down on the Controller VM KB 5228 - NCC Health Check: pcvm_disk_usage_check KB 4872 - LCM: How to cancel an ongoing LCM update operation using the "Stop Update" feature KB 5505 - Long Term Support (LTS) and Short Term Support (STS) Relea
Hi All,We have scheduled to perform a Nutanix Cluster upgrade.Existing Setup:3xDell XC730 XD-12 running on AHV 20170830.453 / AOS(5.15.3). This cluster is hosting around 50 VMs.Upgrade to:4* NX 8235 -G7-4215R -CMIs it possible to add the 4*NX to existing cluster and then remove 3xDell XC servers. We are trying to achieve minimum downtime here without major outage.Per Nutanix documentation https://portal.nutanix.com/page/documents/details/?targetId=Hardware-Admin-Ref-AOS-v5_17%3Ahar-product-mixing-restrictions-r.html , it seems mixing hardware isn't feasible. Could you please confirm if there is any work around possible?Regards,
Below are new knowledge base articles published on the week of May 23-29, 2021.KB 11317 - License Feature Violations or Expired Licenses on Xi-Beam KB 11327 - Calm Blueprint output window block can only be minimized by clicking on the minimize button KB 11339 - X-Ray fails to run the Throughput Scalability test KB 11350 - Recieving "name in body should be at most 40 chars long" when creating a Karbon Cluster KB 11354 - Hit logs from Flow-enabled AHV cluster not forwarded to remote syslog due to regex issue KB 11363 - Lazan service restarts with "GenesisClientError: kRetry: Failed to check if upgrade is in progress, please retry" KB 11367 - Calm: App/VM deployment fails with ERROR - "The resource '<value>' is in use." KB 11395 - Windows VM running VirtIO 1.1.6 may crash with code IRQL_NOT_LESS_OR_EQUALNote: You may need to log in to the Support Portal to view some of these articles.
This is a continuation of the post that I wrote here: https://next.nutanix.com/topic/show?tid=33950&fid=31 Part of this was resolved however we are still running into an issue with some of our VMs in which they get stuck in this state: { "status": { "description": "2016 Template-4/17/19", "state": "ERROR", "execution_context": { "task_uuids": [ "8d0c2110-461b-4252-a223-11016ff8714b" ] }, "message_list": [ { "message": "Edit conflict: please retry change. Entity CAS version mismatch.", "reason": "CONCURRENT_REQUESTS_NOT_ALLOWED" } ]} The only way to clear it is to power the VM down and update the VM within the GUI. However, upon trying to subsequently update the VM by API it results int the same error. Please advise.
Hi All, I am a beginner to Nutanix, just wanted to AHV / AOS! how its function.
The release-api.nutanix.com is not reachable from my prism central and my prism element .I have valid name servers configured in both PC and PE .I got it verified from network team that the traffic is passing by firewall .Can anyone let me know what exact things do i need to check in my name servers so that this URL will be connected from PC and PE ?
Hello Community, I have some doubts about running Nutanix on VMWare with ESXi with Cisco ACI fabric. To start, we have already married ourselves to a Cisco ACI network fabric (two sites connected with 12x10Gbit fiber (120Gbit)). We are using an IPN / Spine / Leaf topology, with APIC clusters in each site. In each DC there will be 24 Nutanix Nodes. All of my question have to do with the best practices of integrating these technologies. 1) Is it necessary to use VMware NSX to reap the benefits of Cisco ACI? 2) Can we just use a simple VMware installation (without NSX) and allow Nutanix full access to the ACI fabric? 3) Whats the best practices related to these three technologies coexisting together? I have found documents on Nutanix/ACI and Nutanix/VMWare, but I cant find anything on using all three together in terms of hierarchy and how to stitch it all together. Any guidance from those who have experience would be greatly appreciated. Thanks, Michael
The reason we deploy SQL Always-on as an enterprise is to have multi-site resiliency for our SQL clusters for operational and business continuity reasons. The issue is that ERA will not register SQL always-on clusters if they are spread between two sites/clusters. From an “API First” solution I would expect differently! I can not introduce risk and pull back DR planning because of the lack of support from ERA. How is the NTX community solving this issue?
Im trying to create a playbook that on X alert posts to teams channel…. is this possible? I have tried a few methods but not having much joy. Method: POST URL: https://outlook.office.com/webhook/blahblahblahlongstring user: blank Pass: blank Request body: I have tried setting json and tried using the built in parameters Alert: Source Entity Name Alert: Alert Name Alert: Creation Time Alert: Severity request headers: blank or Content-Type: application/json so am i just missing something about the syntax? Or is the rest api call for nutanix api’s only? Thanks in advance.
Is it possible for purposes of DR in a cloud solution to move just the VMWare environment?Can you replicate your Production Nutanix VMware environment to a cloud based solution? Say IBM cloud? They say you can but i guess i am having a hard time seeing how that will work? I am really new to the whole cloud solution arena and my company wants to move our DR\Datacenter to cloud? We have business reasons for using IBM cloud but they do not have a nutanix cloud infrastructure but tell us we can bring our VMWare environment.
Below are new knowledge base articles published on the week of May 16-22, 2021.KB 11189 - How to secure the bootloader with a user-defined password KB 11190 - RHEL STIG requirement for actions when audit storage is full KB 11191 - How to configure CVM to use nutanix user on AHV instead of root KB 11192 - How Envoy handles downstream service failures KB 11193 - New users are added to AHV: "nutanix" and "admin" KB 11252 - SSL Certificates and the Secure Gateway Appliance for Frame KB 11312 - AOS upgrade notification shows 5.20.x as Short Term Support (STS) release instead of Long Term Support (LTS) release KB 11321 - Calm VM image upload to vCenter7 fails with error "Unable to retrieve manifest or certificate file." KB 11333 - Adding Nutanix Objects as Secondary Storage with Veritas Enterprise Vault KB 11334 - AHV/ESXi Standalone Foundation deployment failure on HPE DX AMD platforms with Mellanox NIC CX4 or CX5 KB 11346 - Nutanix Files - Inode usage high on FSVM KB 11364 - CVE-2020-13946
AOS 5.20 delivers performance enhancements that build on the breakthroughs in AOS 5.19 and expands on built-in key management capabilities for keeping data encrypted and secure. AOS 5.20 also increases the portability of VMs running on the built-in hypervisor AHV, streamlines advanced management capabilities, and more. A complete list can be found in the Release Notes. Nutanix Insights:Available with all new AOS releases is Nutanix Insights. Most enterprise IT solutions rely on a reactive approach to system maintenance and issue resolution. For example, when a technical issue arises, vendor support teams typically capture detailed system data from the customer and recreate the issue in a separate environment—only then can the actual debugging begin. This approach consumes unnecessary time and resources and ultimately delays resolution. Nutanix simplifies and streamlines this process through two important support services: Nutanix Pulse and Nutanix Remote Diagnostics. When enabled, Puls
On cvm, the cs and cluster start commands are output to the next screen.Hypervisor : ESXI 6.7U3AOS : 5.15.4What's the problem?
I just updated a cluster’s AOS and am now getting the warning:AHV version 20190916.360 is currently installed on host, while minimum compatible AHV version is 20190916.410But I cannot AHV without updating also AOS to 5.19.2 which would change my plan to STS (currently installed 5.15.5.1 LTS)So I have to ignore the warning and waiting for a compatible AOS LTS version, or should I change to 5.19.2 STS?
I have a server that no longer needs one of its 2TB disk. I know how to remove the disk from the VM but, how do I remove that VMDK from the Nutanix Storage container? Is there a way to browse it and then delete it? 2TB is a lot of space that I would like to reclaim.
Which node type does not deploy a Nutanix Controller VM. A bit confused with documentations available some say “Compute only” and other say “All Flash”One more question A vDisk is read by multiple VMs. The cluster creates immutable copies of the vDisk. What are these vDisk copies called.
We noticed that recently in our test cluster all operations take very long to complete. This seems to be due to long lasting “cerberus” tasks being executed with a type “uncharge”. Does anyone know what “cerberus” is doing or how I could speed it up? Whenever we kill the cerberus tasks manually all other queued ‘non cerberus’ tasks start to execute.
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.