Get guidance, share wins, and ensure smooth Nutanix deployments.
Recently active
Leap protects your guest VMs and orchestrates their disaster recovery (DR) to other Nutanix clusters when events causing service disruption occur at the primary availability zone (site). For protection of your guest VMs, protection policies with Asynchronous, NearSync, or Synchronous replication schedules generate and replicate recovery points to other on-prem availability zones (sites). Recovery plans orchestrate DR from the replicated recovery points to other Nutanix clusters at the same or different on-prem sites.Protection policies create a recovery point—and set its expiry time—in every iteration of the specified time period (RPO). For example, the policy creates a recovery point every 1 hour for an RPO schedule of 1 hour. The recovery point expires at its designated expiry time based on the retention policy. If there is a prolonged outage at a site, the Nutanix cluster retains the last recovery point to ensure you do not lose all the recovery points. For NearSync replication (li
Just wondering if i need to convert the thick disk to thin provision disk prior to the migration process from VMware to Nutanix.
The following warning occurred to customer Nutanix.WARN: Cluster cannot tolerate 1 node failure.As a result of the search, we found the following:impact Cluster can not tolerate single node (RF2) or two node (RF3) failure and guarantee available storage to provide data redundancy.The usage warning is 26.93 TiB of the 26.41 TiB recommended usage.The question here is, if one out of four nodes fails,Is it still possible to operate with RF1?Or will the cluster hang up?
Below are the top knowledge base articles for the month of November 2021.KB 7503 - NX Hardware [Memory] – G6, G7 platforms - DIMM Error handling and replacement policy KB 1540 - [AOS Only] What to do when /home partition or /home/nutanix directory on a Controller VM (CVM) is full KB 4409 - LCM: (Life Cycle Manager) Troubleshooting Guide KB 6153 - NCC Health Check: default_password_check and pc_default_password_check KB 3827 - Alert - A130087 - Node Degraded KB 1113 - HDD or SSD disk troubleshooting KB 4158 - Alert - A1104 - PhysicalDiskBad KB 2090 - AHV host networking KB 2473 - NCC Health Check: cvm_memory_usage_check KB 4519 - NCC Health Check: check_ntp KB 10919 - Logrotate does not rotate ikat logs KB 12257 - CVM or PC VM may exceed the thread count limit on the cluster, resulting in services crashing or connection to PE or PC failing with "Server is not reachable" KB 10908 - CVM in a boot loop after a reboot or an upgrade KB 4141 - Alert - A1046 - PowerSupplyDown KB 2463 - NCC Hea
Below are new knowledge base articles published on the week of November 28-December 4, 2021.KB 11493 - NCC Health Check: check_ssl_expiry KB 12377 - Nutanix Files upgrade fails with "Software afs.{version} already exists on the cluster" KB 12402 - Upgrade to SPP 2021.04.0.02 is blocked for certain hardware models KB 12419 - Xi Frame - Users unable to start session with Crowdstrike enabledNote: You may need to log in to the Support Portal to view some of these articles.
The Nutanix AOS/AHV was installed on the DX equipment, and the ISCSI Data Services IP was also set up. The Volume Group tab and create Volume Group are not visible in the Prism - Storage - Table.AOS version is 5.20.xIs Volume Group not supported for DX equipment? Or do I need another setting after AOS 5.20?
Hello Team,Hope your are all doing well We need to size a New Nutanix Cluster for Disaster Recovery purposes of an existing cluster (04 x NX-8155-G7 with AHV).Before that, we need to scale the Existing platform by adding more compute ressources to support additionnal workloads. We were able to do it easily : we used a Scenario in Prism Central and added the workloads and clicked on the “Recommend” Button (please see pic below) : Now, we are not able to do the following :Size the additionnal storage related to retention of local snapshots (for the Production Cluster) Size the Cluster of the DR Site : We used Nutanix collector to collect performance of the existing platform and next we exported the Excel Output file and imported it to Nutanix Sizer. We faced from the following :We are not able to inject the additionnal Workloads for the Production cluster We are not able to add the addtionnal storage related to retention of remote snapshots.Please do you have any idea to correctly size
Hello everybody, I noticed that application consistent VM snapshots of e.g. Windows 10 Pro VMs with NGT enabled and installed in the VM and verified with ‘ncli ngt list’ doesn’t work and results in the following alert… VSS snapshot is not supported for the VM 'TEST-VM', because VSS software is not installed. and … VSS is enabled but VSS software or pre_freeze/post_thaw scripts are not installed on the guest VM(s) TEST-VM protected by TEST-VM. Using a Windows Server 2016 VM with the respective Protection Domain generates application consistent Nutanix VM snapshots. With Windows 10 VMs it won’t work. Verified in 2 different Nutanix clusters. Clusters are equipped with AOS 5.15.4 LTS version. Nowhere I could find a hint, that Windows 10 VMs are not supported. Regards,Didi7
Has anyone used Ansible for Automating Nutanix tasks. Can someone share some examples please.
Below are new knowledge base articles published on the week of November 21-27, 2021.KB 12224 - Unable to shut down the Windows Server 2019 VM from Prism (VM Power operation) when the guest's desktop is locked by the user KB 12332 - VMs which were part of Flow Microsegmentation Security Policies may encounter network connectivity issues if Flow is later disabled KB 12334 - VMs which are part of Flow micosegmentation isolation security policies might encounter network connectivity issues after bulk VM power off/power on events KB 12349 - Prism Central | Reports lists incorrect "Entity Count" when running a report for "specific entities" when NGT rules are applied KB 12358 - Commvault Backups fail with "Unknown VMInfo error" KB 12367 - LCM Pre-check: test_phorest_certs KB 12379 - ERA provisioning error " 'NoneType' object is not subscriptable" KB 12380 - HPDX: NVME disk might not be visible on Prism KB 12393 - Flow Networking (ANC) - Deploying VPN Gateway at a Dark SiteNote: You may need
Hello, I’m facing this issue migrating a Windows VM from vCenter 5.5 to AHV 5.20 with Move 4.2.0. Migration Status “Failed” with Reader Client error: context deadline exceeded error. also in logs I found this errors:in srcagent.logI1124 14:02:28.270928 10 uvmcontroller.go:1092] [VM:'Migration-test1' (moID:vm-3682)] Cancelling context in UVM pre-migration operation. Got Context error: [<nil>]I1124 14:02:28.278034 10 esxvm.go:1205] [VMMoID: VirtualMachine:vm-3682] Prepare access for VME1124 14:03:28.278714 10 diskreaderclient_grpc_impl.go:65] failed to connect to diskreader grpc server running on port: diskreader:8092. Got error [context deadline exceeded]E1124 14:03:28.278811 10 esxvm.go:1168] Failed to connect with Disk Reader Service: [Location="/hermes/go/src/diskreader/client/diskreaderclient_grpc_impl.go:66", Msg="context deadline exceeded"] Reader client error (error=0x1001)E1124 14:03:28.278834 10 esxprovider.go:875] [VM:'Migration-test1' (moID:vm-3
Hi,Does anyone use One-Click for rolling ESXi patching where you apply patch bundles, i.e. many patches at once? Using One-Click for an upgrade or applying a single patch is documented and discussed/known, i.e. when you need to apply a single file. I’m specifically interested in patch bundles. Like just now VMware released a bundle of 10 patches for 6.7. Making 10 rolled updates via One-Click for each patch is impractical. While applying via VUM does it fine, but due to manual CVM shutdown, it can only be done manually host by host, watching Prism for Data Resilience to return, hence taking time.I’ll appreciate if you please share you approach to this.Thank you.
Good afternoon everyone.During the integration with Nutanix AHV, I faced the following case -I need to perform the operation of restoring VM to the snapshot state, my questions are:1. how can I do this with API v3?2. may I restore the state of one VM with the snapshot from another one?What are recommendations and best practices exist to do this with API v3? Thanks.
Using the LCM I updated one of our hosts to the latest firmware from G4G5T6.0 to G4G5T8.0, (BMC 3.64 to 3.65) and since then all I get is a black screen with the words, “boot error”I can get into the bios via IPMI and see the boot device, “sSATA P3: SATADOM-SL 3ME” but it just not booting from it anymore. I also could not find any documents on how to trouble shoot it.Could someone point me in the right direction please, thanks!Note, out of the 6 hosts in our environment, 3 have updated fine already and nothing has changed, so a little confused. I’ve also been updating one host at a time.
You can add a cluster name or virtual facing IP address at any time. To create or modify either value, do the following: Procedure:In the main menu, either click the cluster name at the far left or click the gear icon in the main menu and then select cluster details in the Settings page. The cluster Details window appears. It displays the cluster UUID (Universally unique Identifier), ID and incarnation ID values, along with the cluster name, virtual IP address and iSCSI data services IP address (if configured). The cluster ID remains the same for the life of the cluster, but the incarnation ID is reset (typically to the wall time) each time the cluster is re-initialized. Enter (or update) a name for the cluster in the Cluster Name field.The default name is simply Unnamed. Providing a custom name is optional but recommended. Enter (or update) an IP address for the cluster in the Cluster Virtual IP Address field.A Controller VM runs on each node and has its own IP address, but this f
Nutanix Clusters delivers a hybrid multi cloud platform addressing the pressing need for a single platform that can span private, distributed, and public clouds so that you can manage your traditional and modern applications using a consistent cloud platform.Nutanix Clusters has features such as operational simplicity, seamless application mobility, and cost efficiency needed to run applications in private or multiple public clouds and reduce the operational complexity of migrating, extending, or bursting your applications and data between clouds.Nutanix Clusters extends the simplicity and ease of use of Nutanix hyperconverged infrastructure (HCI) software as well as the full Nutanix stack to public clouds such as AWS.You can modify, update, display, hibernate, resume, and delete Nutanix clusters running on AWS by using the Nutanix Clusters console. Nutanix Clusters console Updates The service that is responsible for provisioning, maintaining underlying infrastructure and making confi
Prism Central (PC) administrators with the "Prism Admin '' role have full access to Karbon and its functionalities. Performing most karbonctl operations requires admin privileges. PC admins that do not have the "Prism Admin '' role (Cluster Admin and Viewer) can only access Karbon to download the kubeconfig and cannot perform any other administrative tasks. See "User Management" in the Prism Web Console Guide for steps on assigning roles.Nutanix requires configuring Karbon users for a directory service in Prism. See Security Management in the Prism Web Console Guide for directions on configuring a directory service.After setting up and testing your cluster, configure role-based access control (RBAC), see Kubernetes documentation for reference. Accessing Locked nodesKarbon protects all nodes in a cluster. You can access nodes in a Kubernetes cluster using an ephemeral certificate, which expires after 24-hours. Perform the following steps to get a certificate. Procedure In the Clusters
Xi Infrastructure ServiceAlong with the infrastructure required to support the recovery of on-prem applications, Xi Cloud Services provides the functionality that you require to run the application in the cloud until you can perform a failback. This functionality, made available as Xi Infrastructure Service, is designed to enable you to address application scaling requirements when the primary availability zone is being recovered from a disastrous event.The configuration tasks described here are specific to Xi Cloud Services. To perform on-prem infrastructure tasks, such as VM and category management, you must continue to use the on-prem Prism Central instance. For information about these tasks, see the Prism Central Guide. Entities MenuThe entities menu provides access to all the entities in Xi Cloud Services. An entity is an object type such as a VM, image, floating IP address, or virtual private cloud (VPC). Virtual Private Cloud Management Floating IP Addresses Management Vir
Creating a Project in CalmYou create a project to map your provider accounts and define user roles to access and use Calm.Ensure that you configure the provider accounts that you want to add to your project. Procedure Click the Projects icon on the left pane. The Projects page appears listing all your existing projects. Click the +Create Project button to create a new project. The Create Window appears. 3. Enter a name for the project in the Project Name field.4. Enter a description for the project in the Description field.5. (Optional) Enter an admin for the project in the Project Admin field.The project automatically adds you as a Project Admin when you create it. You can add more users in the subsequent steps of project configuration.6. Check the Allow Collaboration check box to allow project users to collaboratively manage VMs and applications within the project.The Allow Collaboration check box appears when you add your first user to the project. By default, the Allow Collabor
hi Friends, We have a 3-node Lenovo SR530 with IPMI IPs set. We are trying to install and Setup the Cluster remotely (vpn). Below are the details of IPs,XCC IPs - 172.31.199.x (all 3 nodes of same subnet)AHV IPs - 172.31.199.x (all 3 AHVs IPs of same subnet as XCC)CVM IPs - 192.168.5.x (All CVM IPs with thier default IPs) Problem:I have installed Foundation Applet 5.1 and while running on my laptop, its not able to detect the nodes. Any idea on what to check? The other way suggested in one of the forums was execute the cluster create command in one of CVMs. But all the CVMs have same IPs (192.168.5.2). Should I change the IPs of all CVMs with unique IPs matching with AHV and XCC Subnet IPs and then execute cluster create command? How can I know the CVM IP address from the AHV cliFoundation Applet error messageJava Web Start 11.311.2.11Using JRE version 1.8.0_311-b11 Java HotSpot(TM) Client VMJRE expiration date: 19/2/22 12:00 AMconsole.user.home = C:\Users\KishoreSalipalli-------------
Below are new knowledge base articles published on the week of November 14-20, 2021.KB 11994 - Xi Frame - Audio issues Echo / Delay KB 12075 - Performance benchmarking with Fio on Nutanix KB 12297 - VM Power ON/OFF operations done via the Nutanix AHV Plugin for Citrix takes several minutes to complete KB 12340 - Prism Central 1-click upgrade or pre-upgrade does not show the progress KB 12343 - LCM IVU: Fetch In VM Update (IVU) images fails if http protocol is blocked in the cluster KB 12352 - SNMP stops working due to high CPU usage (100%) for snmpd after a node is rebooted KB 12373 - SNMP user information is deleted after expanding cluster KB 12376 - Dark site LCM updates fail with error "Could not resolve host: download.nutanix.com"Note: You may need to log in to the Support Portal to view some of these articles.
There could be scenarios where LCM (Life Cycle Management) update operation was initiated, and then we need to cancel it in between due to any of the following reasons: 1) LCM operation takes more time than the scheduled maintenance window2) LCM operation was initiated by mistake, and so we want to cancel it soon Starting LCM 2.4.0.3 and later versions, we can now cancel or abort an ongoing LCM update operation using the "Stop Update" feature available on the LCM page via Prism. Whenever this feature is used in LCM, it sets a cancel intent in the Nutanix cluster, which tells LCM to stop the update at the next safe point. The important point is that the update may not be canceled immediately after the cancel intent is invoked since LCM stops it at the next safe point. Hence, we will have to wait until the update reaches the next safe point. This safe point is determined by the LCM framework, which depends on the phase at which the cancel intent was set. Kindly refer to KB-4872 to know m
To replicate entities (protection policies, recovery plans, and recovery points) to different on-prem availability zones (sites) bidirectionally, pair the sites with each other. To replicate entities to different Nutanix clusters at the same site bidirectionally, you need not pair the sites because the primary and the recovery Nutanix clusters are registered to the same site (Prism Central). Without pairing the sites, you cannot perform DR to a different site. To pair an on-prem site with another on-prem site, perform the following procedure at both the sites.Procedure Log on to the Prism Central web console. Click the hamburger icon at the top-left corner of the window. Go to Administration > Availability Zones in the left pane. 3. Click Connect to Availability Zone.Specify the following information in the Connect to Availability Zone window. Availability Zone Type: Select Physical Location from the drop-down list.A physical location is an on-prem availability zone (site). To
Hello, I have a 3 node 1065 G6 cluster running AHV VMs. I currently have six 6TB EXOS drives and want to expand capacity by replacing them with 12TB EXOS drives. Nutanix says that this is not supported, but has anyone replaced drives to expand capacity? I have no money in the budget for another node. Thanks for any info. Jim
Below are new knowledge base articles published on the week of November 7-13, 2021.KB 11979 - HPDX unavailability of SPP upgrade under LCM | Manual SPP Upgrades KB 12252 - NVD - Recovering Calm on AHV in a Different Network Subnet After a DR Event KB 12316 - Era - rc.local service fails on a provisioned Oracle DB on some Linux versions KB 12319 - CVM may fail to reboot on Trimode connected HPDX Gen10 Plus platforms KB 12321 - Files upgrades and FSM requirements KB 12328 - Monitoring cluster recovery process after power outage KB 12331 - SQL monitoring service stops after Prism Central IP is changed KB 12338 - SQL monitoring service may show instances as 'Disabled' KB 12339 - AOS upgrade to 6.0.2 stalls as eth0 interface on the CVM fails to come UP after a reboot KB 12341 - Hyper-V: After upgrading to AOS 6.0.1 or higher the cvm_shutdown -P script fails to shutdown the CVMNote: You may need to log in to the Support Portal to view some of these articles.
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.