Get guidance, share wins, and ensure smooth Nutanix deployments.
Recently active
I am facing an issue with the sizer while generating a budgetary quote. Help please!
Can anyone explain the function of the command “nodetool -h 0 ring” and also the explanation of its output, because I often see nutanix support using this command, i just curious function this command :D
I am trying to execute a remote PS script in a windows server using the playbooks. Have kept the script very simple for now to test the execution, but I get the error “ Failed to execute action with error: Internal Error, Connection reset by remote peer.” every time its executed.Execution Start Time: 01/06/21, 5:21:32 PM Execution End Time: 01/06/21, 5:21:32 PM Result: FAILEDInputUsername: serveradminPath to Script: E:\NutanixPS\Nutanix-CVM-Post-reboot-checks.ps1Password: ********IP Address/Hostname: 10.32.135.100HTTPS: True Has anyone faced a similar issue and what is the fix?
Hi I deploy a 3 nodes cluster with ESXi 7.0 as the hypervisor successfully on AOS 5.15.3.But when I run the diagnostic.py to check the cluster performance. It always failed to boot the diagnostic VM. I tried 3 times and all is same. from the log, seems the CVM cannot reache the diagnostic VM when check the boot status.I login to vCenter and check that there is no IP address for this VM in the summary page. And I can login the diagnostic VM from the VM console, I can check the ip address which is 192.168.5.253, but I cannot ping 192.168.5.254 and 192.168.5.1. also cannot ping 192.168.5.253 from CVM.From this information, I think the CVM cannot reache the diagnostic VM and timeout, so the diagnostic.py will fail after timeout. Is there any idear for this test? I can run thie script successfully in ESXi 6.7.
Hello. i Have a 4 nodes Lenovo HX3320. When i want to install AOS it fails. Foundation 4.6AOS 5.15.4 I updated bmc firmware and reset it. But it didnt work.Can you help me with this problemhere is the log 2020-12-10 12:38:39,191Z DEBUG Setting state of <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0> from PENDING to RUNNING2020-12-10 12:38:39,197Z INFO Running <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0>2020-12-10 12:41:01,398Z DEBUG Cache HIT: key(<function common_validations at 0x03D71230>_()_{'global_config': <foundation.config_manager.GlobalConfig object at 0x054D0E90>})2020-12-10 12:41:01,405Z DEBUG Setting state of <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0> from RUNNING to FINISHED2020-12-10 12:41:01,410Z INFO Completed <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0>2020-12-10 12:41:01,415Z DEBUG Setting state of <GetNosVersion(<NodeConfig(192.1
I have 20 clusters under management of our Prism Central and would like to know the versions of NCC installed in each. Is there a way to singularly collect this info across all PE clusters? Perhaps in nCLI even?
So guys I’ve got a flexlm license server running Windows Server 2019 that I am trying to use Move to migrate into my AHV cluster. I tried migrating it once and it blew up the license server since I guess it builds a unique machine ID using a disk and perhaps CPU ID. Any idea on how to clone this in Move so when it comes across FlexLM and other license daemons are none the wiser?
This article contains IPMI commands for checking and setting interfaces to dedicated or shared mode. For example, after a BMC upgrade, the IPMI might not be accessible. So, you need to verify and change the interfaces to dedicated or shared mode. Note: To run ipmitool commands on an ESXI host, prefix all commands with a forward slash (/). Note: To run ipmitool commands from a remote system such as the CVM (Controller VM), add the "-I lanplus", "-H <IPMI IP>", "-U <username>" and "-P <password>" parameters to the ipmitool command. For example: nutanix@cvm$ ipmitool -I lanplus –H x.x.x.x –U ADMIN –P <password> <command> Quanta Platform Use these commands for an NX-3400 (Quanta) platform. All commands are executed dynamically and a restart is not required. Check the status. [root@host]# ipmitool raw 0x0c 0x02 0x01 0xff 0 0 An output similar to the following is displayed.1100 :00 - Shared port 1101 :01 - D
In some cases, you might have to permanently remove a physical node / host from a Nutanix cluster. There are two scenarios in node removal. Permanently Removing an online node Removing an offline / not-responsive node in a 4-node cluster, at least 30% free space must be available to avoid filling any disk beyond 95%. You cannot remove nodes from a 3-node cluster because a minimum of three Zeus nodes are required. Some Points to consider before initiating node removal: Sufficient Disk space available on other nodes in the cluster User Virtual Machine relocation (if required) Any software upgrade should not be running Checklist on verifying cluster health status Data resiliency is “OK” (green) in Prism Run a complete “ncc report” either from prism or CVM cli: ncc health_checks run_all Depending on the size of data, node removal can be lengthy process, which involves relocating data from the node to other healthy nodes in the cluster. Node removal also remo
Hi,I am implementing a Nutanix Cluster with Lenovo HX5520, when I register the Prism in vCenter the process appears as completed but I cannot create virtual machines by Prism Element and when running the NCC I return communication errors with the vCenter is not stabilized. I did the communication test of the CVMs with the vCenter on ports 80 and 443 and the connection worked successfully. when checking by cli I see that the connection is ok, but it does not provide me with the settings of vcenter follows the difference from another implementation any idea what it might be?
Below are new knowledge base articles published on the week of December 27, 2020-January 2, 2021.KB 10019 - Alert - A110024 - AwsDefaultUVMSecurityGroupNotFound KB 10507 - Nutanix Move | VM migration fails if Hyper-V VM has Fibre Channel disk controller attached KB 10528 - Era - Operation failed with Internal Error after changing Cluster Account Password KB 10529 - [ Karbon ] How to configure email alerts for a Karbon kubernetes cluster KB 10530 - NCC-4.0.0 : Health Server logs might fail to rotate and fill up /home partitionNote: You may need to log in to the Support Portal to view some of these articles.
Below are the top knowledge base articles for the month of December 2020.KB 7503 - NX Hardware [Memory] – G6, G7 platforms - DIMM Error handling and replacement policy KB 10475 - LCM 2.4 inventory failure - [SSL: UNKNOWN_PROTOCOL] unknown protocol (_ssl.c:618). Not fetching available versions for module KB 4141 - Alert - A1046 - PowerSupplyDown KB 1540 - What to do when /home partition or /home/nutanix directory on a Controller VM is full KB 1113 - HDD/SSD Troubleshooting KB 4409 - LCM: (LifeCycle Manager) Troubleshooting Guide KB 4158 - Alert - A1104 - PhysicalDiskBad KB 2090 - AHV host networking KB 4519 - NCC Health Check: check_ntp KB 2473 - NCC Health Check: cvm_memory_usage_check KB 4116 - NX Hardware [Memory] – Alert - A1187, A1188 - ECCErrorsLast1Day, ECCErrorsLast10Days KB 4273 - NCC Health Check: aged_third_party_backup_snapshot_check and aged_entity_centric_third_party_backup_snapshot_check KB 6945 - How Upgrades Work at Nutanix KB 1863 - NCC Health Check: sufficient_disk_s
What is happening to my 2 node cluster during a failover or an upgrade?What does a recovery process look like after a node failure?If you are wondering the above, we have the answer for you! You can monitor the progress of your 2 node cluster in these situations through Prism Element.To monitor node recovery progress after failover: Registering a witness is highly recommended to help the cluster handle the failover situation automatically and gracefully. Stand-Alone mode: A failed node would trigger cluster to transition into stand-alone mode during which the following occurs: Failed node is detached from metadata ring. Auto rebuild is in progress. Surviving node continues to serve the data. Heartbeat: Surviving node continuously pings its peer. As soon as it gets a successful reply from its peer, clock starts to ensure that the pings are continuous for the next 15 minutes. If a ping fails after a successful ping, the timer will be reset. Prism Element Home page shows Critical
Prism Central includes machine-learning capabilities that analyze resource usage over time and provide tools to monitor resource consumption, identify abnormal behavior, and guide resource planning. These tools include VM "right sizing" where VMs are analyzed and those that exhibit inefficient profiles are identified. Anomaly detection to record when performance or resource usage is outside an expected range based on learned VM baseline behavior. "Smart" alerts that trigger when specified anomalies are recorded. Reports that summarize cluster efficiency. VM Right SizingIt is useful to look at the profile of your VMs when analyzing problems in a cluster or assessing future resource needs. This can help you identify VMs that are not optimally configured such as ones that consume too many resources, are constrained, are over provisioned, or are inactive.Anomaly Detection:The right sizing feature identifies inefficient VMs that fit one of the profiles described as below: Bully VM : A
Nutanix takes a holistic approach to security with a secure platform, extensive automation, and a robust partner ecosystem. The Nutanix security development life cycle (SecDL) integrates security into every step of product development, rather than applying it as an afterthought. The SecDL is a foundational part of product design. The strong pervasive culture and processes built around security harden the Enterprise Cloud Platform and eliminate zero-day vulnerabilities. Efficient one-click operations and self-healing security models easily enable automation to maintain security in an always-on hyperconverged solution.Since traditional manual configuration and checks cannot keep up with the ever-growing list of security requirements, Nutanix conforms to RHEL 7 Security Technical Implementation Guides (STIGs) that use machine-readable code to automate compliance against rigorous common standards. With Nutanix Security Configuration Management Automation (SCMA), you can quickly and continu
Below are new knowledge base articles published on the week of December 20-26, 2020.KB 10328 - Windows VM on AHV with Nutanix VirtIO Unable to Read "Physical Disk Serial Number" Intermittently KB 10349 - DHCP-client startup impacting Windows VM guest services KB 10415 - NX-8170-G7 imaging fails with Foundation < 4.5.4 KB 10446 - Cannot provision node due to AWS Quota exceeded issue. Quota type cpu. KB 10456 - ESXi 6.5 failure when imaging with Foundation 4.5.4.2 KB 10474 - Objects - Manual steps to configure emails for Objects Alerts KB 10487 - Unable to create thick provision disks from Nuranix NFS datastore on VMware due to nfs-vaai plugin missing on ESXi hosts KB 10508 - Rack aware settings not working when PE launched from PCNote: You may need to log in to the Support Portal to view some of these articles.
Here I discuss the effects of Enabling or Disabling Deduplication on a container even if the Container has data already written to it. The benefit of Compression and Fingerprinting+Deduplication is to hold more data in the container, by reducing the stored size and avoiding duplicate data, respectively.Nutanix’s intelligent selection of dedupable candidates prevents deduplication being performed where the benefit would be low. Deduplication Best Practices: Enable deduplication Do not enable deduplication Full clones Physical-to-virtual (P2V) migration Persistent desktops Linked clones or Nutanix VAAI clones: Duplicate data is managed efficiently by DSF so deduplication has no additional benefit Server workloads: Redundant data is minimal so may not see significant benefit from deduplication Enabling Dedupe:Fingerprinting is method of creating signatures of the data in Metadata. Fingerprint-on-write (Cache-Ti
I am moving virtual machines with Move 3.6.2 from Hyper-V to an AHV CE cluster. Everything goes well, but after the cutover the AHV VM is stuck at the ‘Press F2 for EFI boot manager’ screen and cpu usage around 52%. Pressing FN+F2, F2, CTRL+F2, ALT+F2 or other combinations have no effect. Please advise. Hyper-V host is Win2019, VM is 2019 with boot from bootmgfw.efi and secure boot disabled. Also tried with Edge, Chrome and IE, same result, stuck at ‘Press F2 for EFI boot manager’ screen.
Have you guys been utilizing the Analysis charts feature effectively? It gives you the ability to create charts that can monitor a variety of performance metrics over a week, month, or a custom time range. Here are a few ideas on how we resolved some issues seen in the field. Memory Usage (%) - Create a chart to track memory consumption or one or multiple VM’s over a time interval Hypervisor CPU Usage (%) - Measure the CPU usage of one or more hosts over time and gain a better understanding of resource constraints if any Storage Controller Latency - This is particularly useful if you suspect performance issues. Creating charts ranging to a few weeks back gives you a comparative analysis and a benchmark for the current latency observed Storage Container Usage - A classic use case for this is when multiple VM’s have been migrated out of the cluster over many days but you suspect the storage space has not been reclaimed Replication Bandwidth - Transmitted - If there are
This reference covers the v1 Nutanix API. The complete reference for the v2 Nutanix API, including code samples in multiple languages, and tutorials are available at http://developer.nutanix.com/ Users Get Logged In Users DetailsGET /users/logged_in_users Get Logged In Details of a userGET /users/logged_in_users/{userName} Get Logged In Users DetailsGET /users/logged_in_users Get Logged In Details of a userGET /users/logged_in_users/{userName} Get Logged In Users DetailsGET /users/logged_in_users path /users/logged_in_users method GET nickname getAllLoggedInUsersInfo type get.base.EntityCollection<get.dto.auth.UserDTO> Property Type Format entities array errorInfo get.base.ErrorInfo metadata get.base.Metadata Get Logged In Details of a userGET /users/logged_in_users/{userName} path /users/logged_in_users/{userName} metho
Below are new knowledge base articles published on the week of December 13-19, 2020.KB 8562 - NCC Health Check: robo_witness_configured_check KB 8563 - NCC Health Check: robo_witness_state_check KB 8565 - NCC Health Check: robo_cluster_witness_sync_check KB 9271 - NCC Health Check: ahv_fs_integrity_check KB 9472 - NCC Health Check: category_protected_vms_multiple_fault_domain_check KB 9525 - Alert - A200330 - Prism Central home partition expansion check KB 9713 - Alert - A130340 - MetroConnectivityUnstable KB 9716 - NCC Health Check: stale_synchronous_replication_parameters_check KB 9845 - NCC Health Check :- "file_server_cvm_config_check" KB 9988 - Pre-Upgrade Check: test_if_expand_cluster_is_not_in_progress KB 10000 - NCC Health Check: objects_deployed_on_unsupported_pe KB 10248 - Alert - A130340 - Cross-container disk migration task is paused. KB 10323 - Move VMs from Protection Domain to Category for Leap KB 10339 - Skipping application consistent snapshot for VM with NVMe disks KB
Hello i have a 4 Lenovo HX 3320 nodes. I updated all bmc on this servers. When i am installing new aos throw Foundation 4.6 its fails with this error Hele is log 2020-12-10 12:38:39,191Z DEBUG Setting state of <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0> from PENDING to RUNNING2020-12-10 12:38:39,197Z INFO Running <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0>2020-12-10 12:41:01,398Z DEBUG Cache HIT: key(<function common_validations at 0x03D71230>_()_{'global_config': <foundation.config_manager.GlobalConfig object at 0x054D0E90>})2020-12-10 12:41:01,405Z DEBUG Setting state of <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0> from RUNNING to FINISHED2020-12-10 12:41:01,410Z INFO Completed <ImagingStepValidation(<NodeConfig(192.168.1.22) @e2d0>) @e5f0>2020-12-10 12:41:01,415Z DEBUG Setting state of <GetNosVersion(<NodeConfig(192.168.1.22) @e2d0>) @e550> from PENDIN
LCM (Life Cycle Manager) is a tool provided by Nutanix to upgrade firmware and software across a Nutanix environment. Depending upon the type of firmware that needs to be upgraded, a reboot of a node may be required.While reviewing the inventory of software/firmware that needs to be upgraded, you may not know which upgrades require a reboot of a node (or not). This information might be crucial in determining the impact and precautions that need to be taken in planning for the upgrade event (change windows, time allocated, etc.).KB 6107 details which upgrades require a node or CVM reboot and which ones do not. Navigate to the
Many users are unaware that there are additional (beyond what is presented via the Prism user-interface) security parameters that can be employed on AHV hosts to increase the overall security of them. These security parameters are configured via Nutanix Command-Line Interface (NCLI) and include the following: Advanced Intrusion Detection Environment (AIDE) - a file and directory integrity checker High Strength Password Enforcement - configure the maximum and minimum number of characters the password must contain along with number of passwords retained in history to prevent repeated use Core Dumps - the recorded state of the working memory for a process is dumped to a file if the process ever crashes Login Banner - display a customized messages when user login to a node More information regarding these parameters, including the procedures to enable/disable them, can be found within the Hardening AHV section of the Nutanix Security Guide. Also to note, there are similar parameters
SAP helps customers migrate from traditional relational databases to their in-memory SAP HANA database to gain more agility in their business processes. Many SAP customers are searching for ways to deploy SAP HANA in an efficient, simple way that minimizes risk while preserving the benefits of an agile platform. Nutanix provides such an option. The native Nutanix hypervisor, AHV, and Nutanix enterprise cloud OS software are certified for production SAP HANA deployments. HCI for SAP HANA CertificationThe certification has two primary segments. 1. As the first step, a platform vendor (Nutanix, in this case) must validate their platform, which consists of a hypervisor and an HCI component.2. In a second step, the hardware OEM must certify a suggested configuration through some additional HCI-related tests. When both parts of the validation are complete, the solution is certified and listed in the HCI for SAP HANA category on the SAP website. The hardware OEM is then responsible for selli
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.