Get guidance, share wins, and ensure smooth Nutanix deployments.
Recently active
Hi i am tring to crate docker containers via my windows installation can someone please direct me where do i put the AHV drivers ? i am getting this msg while i am tring to use the driver Driver not found. Do you have the plugin binary accessible in your PATH?
We have had two instances where a node detected/reported a fault event and reset rebooting vm on each occasion. There seems no reason for this to have happened. Details Host 192.168.xx.x4 appears to have failed. High Availability is restarting VMs on hosts throughout the cluster. 08-17-16, 02:01:41am Host 192.168.xx.x4 appears to have failed. High Availability is restarting VMs on hosts throughout the cluster.08-11-16, 07:19:48am We updated the AHV and NCC and since had a repeat last night from the first instance last week Is there a potential hw fault with the host that has not yet been detected or checked?
We moved a NX3000 to a new environment but one of the nodes died. So we had to reimage that node. The other two where still up and running. But after a power-failure on the rack, the two nodes rebooted but won't come up anymore. Genesis showing some failures:2016-08-18 14:38:40 INFO node_manager.py:3492 Svm has configured ip 10.160.35.124 and device eth0 has ip 10.160.35.1242016-08-18 14:38:43 INFO node_manager.py:3542 Setting up key based SSH access to host hypervisor for the first time...2016-08-18 14:38:43 INFO hypervisor_ssh.py:32 Trying to access hypervisor with provided key...2016-08-18 14:38:46 INFO hypervisor_ssh.py:40 Failed.2016-08-18 14:38:46 INFO hypervisor_ssh.py:44 Trying to access hypervisor with provided password...2016-08-18 14:38:49 INFO hypervisor_ssh.py:52 Failed2016-08-18 14:38:49 ERROR node_manager.py:3547 Failed to set up key based SSH access to hypervisor, most likely because we do not have the correct password cached. Please run fix_host_ssh command manually to
Hey, We're in the process of getting an exchange server virtualized and according to Nutanix best practice, we should create a container with EC-X enabled. When I go to create the container in the PRISM gui, there's no option to enable erasure coding. This contradicts documentation where it shows the option under "advanced options" but on my end, it's not appearing.
We had an issue when the cluster was first setup which involved the hdd. These were moved and replaced and now we see an error reported. It was initially cleared and now resurfaced, this despite a AHV update from 4.6.2 to 4.6.3 and also update on NCC to 2.6.6 Error Detailed information for disk_online_check:Node 192.168.xx.x4:FAIL: Disk '/home/nutanix/data/stargate-storage/disks/S460AATC' failed on node with ip u'192.168.xx.x3'. Disk '/home/nutanix/data/stargate-storage/disks/S460AATC' failed on node with ip u'192.168.xx.x4'.Refer to KB 1536 for details on disk_online_check or Recheck with: ncc health_checks hardware_checks disk_checks disk_online_check smartctl on the hdd shows no errors There seems tio be an issue with retaining or updating info/config on the cluster nodes
Hi, I apologize in advance for the long post !!!! From the information I've reviewed, the Xpress Models (I'm looking at Lenovo's offering) will support a maximum of 4 nodes. In addition there is no Protection Domain. Let's pretend that I'm a cloud provider for multiple customers, and I host their (Domain/File/Print/Email/Sql) servers in my 4 node domain. I want to use an (Acronis/Storagecraft) in VM backup software. These products store their backup images to a NAS/Share. If I setup a second 3 node Xpress Model that is storage "heavy" running Acropolis File Services, as a NAS/Share destination for my backup images, is there any restriction that you can think of that would prevent/limit me from doing this ? Right now we use StorageCraft and save to a Windows Storage NAS that has RAID 10 and I'm concerned that writing to will not be able to handle the I/O? I'm guessing that the 3 node Xpress AFS will easily handle intensive I/O writes. On weekends, the Full backup runs a
https://next.nutanix.com/t5/Installation-Configuration/basic-capacity-calculate-and-RF-question/m-p/10728/highlight/true#M1314 Hi All, referring to the above thread - where Jon is mentioning http://designbrewz.com/ for capacity calculation. does this calculator takes into account Replication factors and the "Extent Store HDD" is what the user will see when configuring storage pool > container ? thank you in advance kind regards
Need details how to store passwords as a secure string and use in a script for access to multiple Prism sites. I have a script to collect storage data from multiple Nutanix sites via Prism. Now I need to remove the plain text passwords. I have located references but only find "NOTE: for security reasons we should store our passwords as a secure string, by declaring these as variables before starting PowerShell." Can anyone provide the steps required to make this work?
All, For those using Palo Alto VMs and AHV please be aware that there is a major bug when upgrading to PANOS 7.1.3 that will corrupt the firewall interfaces randomly and so far there is no fix for it, not even downgrading back to previous versions, in my case was PANOS 7.0.5h2. I am currently working with Palo Alto engineering into fixing the issue, I will update once the case is resolved. For the time being I would highly recommend to stay away from the 7.1 series of upgrades. Regads Juan
Hi, I'm doing a PoC for a client, and I can't browse to Prism with 1 of the CVM IP address. The rest are OK. The cluster is up and running with no error. seems the httpd service unable to start. how do i fix this? below is the error snapshot. Please help. Thanks in advance.
Hi folks, Do you guys have ever think about the question. I read some guides and confused now. 1) node memory number: It seems every type of node has different supported memory number configuration: For NX-1064-G4, It shows only support 16x dimm ! What happens if I install any number, i.g 1, 2, 3,4,5,6 or any number 2) node memory installed slot location: It also says there`s an fixed slot order for different number dimm scenario. what happens if I install dimm not according to the guide ? 3) memory type: LR or R DIMM it also says there`s two type dimm, How to confirm which type is supported by specific type of node ? P.S my environment I have a node NX-1065-G4 with default 64g(4x16g installed in 1a,1b,1e,1f) and now I add two additional 16g dimm into slot 1g,1h. It work fine till now. but my memory number confuration is not supported by the guide ! the guide says NX-1065-G4 only support full slot (16x) dimm configuration ! so the guide has no tips about the slot order too while numb
Hi, Does anyone have a step-by-step guide on creating a Windows cluster on ESX on Nutanix? I've seen various examples online of using ISCI, SMB etc file shares but I'm not entirely sure on what to do. I've got the cluster as far as the file share witness working but how do I then configure the disks that make up the cluster resources? I will eventually be clustering SQL 2012 on this. Cheers, Steve
I occasionally get errors on network validation when using the offline Foundation installer. Usually this is resolved by trying another range- occasinally not. I never have this issue when using the standalone version of Foundation. I am using a completely flat switch (which has had many successful installs), additionally this is with Foundation 3.2.2 and Java 7.45. I can log into the hosts/cvms and get successful pings from all of the IP's I'm intending on using but offline Foundation is adamant that some are unreachable. I solve this by using standalone (which can be a little cumbersome on the road). Just wondering if anyone else is experiencing this as well.
Hi everyone, I imaged a nutanix cluster these days, and found strange phenomena. I found many error key words in foundation node log, but imaging showed successful finally. I clicked "retry imaging failed nodes with last config" many times, after some node failed, and it succeed. I`d like to know why it failed and succeed after just retry. also the foundation process tab showed 3 node failed at 78%, and the node log showed it`s finished ! and after retry, it`s 100%. I`m very worried about if it`s truly successful and there`s potential risks. p.s my environment: Foundation 3.1.1, AOS 4.6.1, ESXI5.5U3D,NX-1065-G4
Hello Folks, NX-1065 has only two 10G port, and two 1G port. and I have only one pair of 10G switch, and one pair of 1G switch. In my opinion, I`d like to make ESXI port(mgmt, vmotion,nfs vmkernel), and CVM port(prism mgmt, cluster mirror ) run on 10G switch, and configure 10g switch uplink to other network which I can connect to ESXI MGMT and prism. make vm data port connect to separate 1g vswitch. but network guy told me the 10g switch can`t uplink to externel network, so I`d like to know Can I move the ESXI port (mgmt) and CVM mgmt(VIP) to 1G vswtich, and create addtional vmotion,nfs kernel port connected to 10G vswtich, also leave CVM cluster mirror interface connected to 10G vswtich. networking design is the key things.
Hi guys, two basic questions: 1)can cvm cluster and ESXI data use the same 10g switch with no performance penelty ? (I have only one pair of 10g switch, and NX1065 one dual port card. As we know, cvm cluster must use 10g network, but I`d like also to use 10g for ESXI data network. ) 2) generally ESXI mgmt(usually,vmkernel port for NFS access) is on the same 10g link with cvm cluster(also CVM prism mgmt), If the user`d like to have a separate uplink or subnet for ESXI MGMT, is it ok ? I`m worried about the performance problem. thanks in advance !
Hi, We have 3 Hyper-V hosts those local admin passwords got expired and i changed them (and they are enabled). Now WSMan Connectivity Check fails in Prism. I have run these lines in CVM for all three hosts but it won't help: ncli managementserver edit name=hostIP password='newhostpass'ncli host edit id=cryptical-host-id hypervisor-password='newhostpass' Any ideas? Our nutanix version is 4.6.1.1.
Hi Folks, For Nutanix product, Is it a best practice to have the latest software version (e.g AOS,NCC) installed, No matter first-time installation, or upgrade. Some venders may advise just install the slightly lower version then the newest. My real user case: we installed NX 2 months ago, and plan to put it into produciton these days, but two new AOS version has been available on Portal.nutanix.com, So Can I update AOS to the newest one ?
One of the CVM is not responding. On console below screen shots are there. What is the soln?
We are fairly new to Nutanix. We have two seperate clusters in seperate datacenters. Our goal is to be able to move workloads between clusters based on needs and resources. We are using VMware 5.5 and everything seems to work as planned with the exception of the moving from cluster to cluster. We have cluster A and cluster B. We are using metro availability to sync the data from cluster A to cluster B at this time and that seems to work correctly. When I attempt to perform a cold migration from cluster A to cluster B I get the following error:Relocate virtual machineArtemisCloneFile/vmfs/volumes/ec936530-956aa-9ac/ArtemisClone/ArtemisClone.-vmdk was not found cold migration from cluster B to cluster A works fine. the whitelists are identical on both clusters. VMWare isn't really much help as their reply is to just browse the datastore and add it to inventory on a server in cluster B
Hi, We have applied license to Nutanix cluster running on vSphere. Then we created cluster with AHV from scratch without reclaiming it from Nutanix portal. How to regenerate license for newly created license? Thanks in advance. Vivek
Hi, We have setup in pri production stage. Cluster is having 2 x 800 GB disk & 2 x 6 TB disk on each node. Cluster is having 3 nodes. We removed single disk from one host & after 10 min. we reinserted disk in same slot. Now cluster is showing 11 disk & not allowing us to format disk which is reinserted. How to add this disk to current cluster? Thanks in advance. Vivek
Hi Team, I executed the NCC health checks run_all on two clusters and got the following message (at the end of the run_all script): Detailed information for sar_stats_threshold_check:ERR : Execution terminated by exception IndexError('list index out of range',):Traceback (most recent call last): File "/home/hudsonb/workspace/workspace/ncc-2.0.2-stable_release/builds/build-ncc-2.0.2-stable-release/ncc-python-tree/bdist.linux-x86_64/egg/ncc/ncc_utils/plugin_utils.py", line 128, in handle_exceptions result = fn() File "/home/hudsonb/workspace/workspace/ncc-2.0.2-stable_release/builds/build-ncc-2.0.2-stable-release/ncc-python-tree/bdist.linux-x86_64/egg/ncc/plugins/base_plugin.py", line 740, in result = putils.handle_exceptions(lambda : check(*check_args), cls.canvas) File "/home/hudsonb/workspace/workspace/ncc-2.0.2-stable_release/builds/build-ncc-2.0.2-stable-release/ncc-python-tree/bdist.linux-x86_64/egg/ncc/plugins/health_checks/sar_checks.py", line 358, in check_threshol
Hi all, I have a upcoming appliance installation and it might be necessary that we need to downgrade the current NOS version. Is there an out-of-the box approach / method available that could be used to downgrade an existing cluster? Or maybe KB-articles that could help me regarding this issue? Thank you and best regards, Andreas
Hello Folks, I`m busy installing a 6 node cluster, some errors or doubt comes out along. my environment : 3 block, 6 nodes (2 nodes/block) AOS 4.6.2 一. When I create a new container, make the Enable Erasure coding button checked, error pops up"enable EC-X will break the current block awareness !" seems mean ECX will disable block level awareness ? 二. As many of us know, if need locate the physical slot location of a disk, you can simply led on that drive. but there is possibility led on function inaccurate or disabled, in this scenario, you must locate the disk by disk id rule, e.g shelf id.slot/bay ID, but the DISK ID in prism hardware section seems ramdom/ruleless number, I can`t locate physical location for a disk. 三. When I do the test about removing a node from nutanix cluster, A doubt appears: Do I need or is it the best practice to execute storage vmotion(other then only vmotion vm`s ) before removing the node ? I think the diffecrence of two methods is who is the data mov
Already have an account? Login
No account yet? Create an account
Enter your E-mail address. We'll send you an e-mail with instructions to reset your password.