<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>vCenter Server Appliance on Netdata</title><link>https://www.netdata.cloud/tags/vcenter-server-appliance/</link><description>Recent content in vCenter Server Appliance on Netdata</description><generator>Hugo</generator><language>en-us</language><atom:link href="https://www.netdata.cloud/tags/vcenter-server-appliance/index.xml" rel="self" type="application/rss+xml"/><item><title>vCenter '503 Service Unavailable': the vSphere Client will not load</title><link>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vcenter-503-service-unavailable/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vcenter-503-service-unavailable/</guid><description>&lt;p&gt;A 503 from the vSphere Client means the reverse HTTP proxy (&lt;code&gt;rhttpproxy&lt;/code&gt;) accepted the TLS connection but could not reach the backend it routes to. The proxy itself is healthy. One of its dependents, typically &lt;code&gt;vpxd&lt;/code&gt;, &lt;code&gt;vmware-vapi-endpoint&lt;/code&gt;, &lt;code&gt;vmware-stsd&lt;/code&gt; (STS), or the HTML5 client backend (&lt;code&gt;vsphere-ui&lt;/code&gt;), is stopped, still starting, or crash-looping. The error string often reads &amp;ldquo;Initialization of one of the components failed.&amp;rdquo;&lt;/p&gt;&#10;&lt;p&gt;Running VMs are unaffected. The hypervisor plane keeps scheduling and serving I/O. What you lose is the management plane: DRS stops rebalancing, HA cannot be reconfigured, vMotion orchestration is gone, and provisioning is blocked. The urgency is operational visibility and control, not workload survival.&lt;/p&gt;</description></item><item><title>vCenter 'Cannot complete login due to an incorrect user name or password': SSO failures</title><link>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vcenter-cannot-login/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vcenter-cannot-login/</guid><description>&lt;p&gt;The &amp;ldquo;Cannot complete login due to an incorrect user name or password&amp;rdquo; string is the exact message operators see in the vSphere Client, in PowerCLI sessions, and in API responses when SSO authentication fails. The text is misleading: the cause is rarely a typo. For a single user it is usually a credential or permission problem. For every account at once it is an SSO/STS infrastructure failure.&lt;/p&gt;&#10;&lt;p&gt;The first triage question is scope: does the local SSO administrator account (&lt;code&gt;administrator@vsphere.local&lt;/code&gt;) still work? If yes, the STS signing certificate and token service are healthy, and the problem is in an identity source (AD/LDAP) or a service account. If &lt;code&gt;administrator@vsphere.local&lt;/code&gt; also fails, the STS infrastructure itself is broken: expired STS signing certificate, clock skew rejecting SAML tokens, or STS memory pressure.&lt;/p&gt;</description></item><item><title>vCenter /storage/db full: vPostgres stops and the whole management plane dies</title><link>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vcenter-storage-db-full/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vcenter-storage-db-full/</guid><description>&lt;p&gt;The vSphere Client returns 503 Service Unavailable. PowerCLI sessions hang and time out. DRS has stopped evaluating, vMotion orchestration is gone, and provisioning fails. Running VMs on the ESXi hosts continue to operate, but the management plane is gone.&lt;/p&gt;&#10;&lt;p&gt;The root cause is almost certainly the &lt;code&gt;/storage/db&lt;/code&gt; partition on the vCenter Server Appliance (VCSA). This is where vPostgres keeps its data files. At 95% utilization on any partition, VMware automatically shuts down &lt;code&gt;vmware-vpxd&lt;/code&gt; to protect the database from corruption. At 100%, vPostgres cannot extend a data file or write a WAL record and crashes. Once vPostgres is down, &lt;code&gt;vpxd&lt;/code&gt; has no database and cannot restart.&lt;/p&gt;</description></item><item><title>vCenter /storage/seat full: stats, events, alarms, and tasks outgrowing their partition</title><link>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vcenter-storage-seat-full/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vcenter-storage-seat-full/</guid><description>&lt;p&gt;The &lt;code&gt;/storage/seat&lt;/code&gt; partition on the vCenter Server Appliance (VCSA) holds the vPostgres tables for Stats, Events, Alarms, and Tasks: &lt;code&gt;vpx_event&lt;/code&gt;, &lt;code&gt;vpx_event_arg&lt;/code&gt;, &lt;code&gt;vpx_task&lt;/code&gt;, and the &lt;code&gt;vpxd_hist_stat*&lt;/code&gt; rollup tables. In modern VCSA it is a dedicated mount, so it can fill while &lt;code&gt;/storage/db&lt;/code&gt;, &lt;code&gt;/storage/log&lt;/code&gt;, and &lt;code&gt;/&lt;/code&gt; all show healthy utilization. Operators checking only &lt;code&gt;/&lt;/code&gt; or the VAMI dashboard&amp;rsquo;s &amp;ldquo;VCDB&amp;rdquo; usage will miss it until vpxd refuses to start.&lt;/p&gt;&#10;&lt;p&gt;When &lt;code&gt;/storage/seat&lt;/code&gt; crosses 95% utilization, vpxd refuses to come up to avoid database corruption. Without vCenter: DRS stops scheduling, vMotion is gone, HA cannot be reconfigured, no provisioning, no management operations. VMs on ESXi hosts keep running because the data plane is independent of vCenter, but everything that touches vCenter is broken.&lt;/p&gt;</description></item><item><title>vCenter certificate expired: the STS signing cert outage nobody saw coming</title><link>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vcenter-certificate-expired/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vcenter-certificate-expired/</guid><description>&lt;p&gt;vCenter is down. Not &amp;ldquo;slow&amp;rdquo; or &amp;ldquo;degraded.&amp;rdquo; Down. The vSphere Client shows a white screen or a 503. PowerCLI sessions fail to connect. API calls return authentication errors. ESXi hosts show as disconnected in bulk. Every integration that depends on vCenter (NSX, vRA, SRM, backup products) has lost connectivity simultaneously. VMs on the hosts are still running, but you cannot manage, migrate, or orchestrate anything.&lt;/p&gt;&#10;&lt;p&gt;You check the browser certificate on the vCenter URL. It looks fine. Months left. You check NTP. Synchronized. You check disk space. Plenty. Nothing in your standard monitoring explains why the entire management plane went dark at once.&lt;/p&gt;</description></item><item><title>vCenter Server Appliance Monitoring</title><link>https://www.netdata.cloud/monitoring-101/vcsa-monitoring/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/monitoring-101/vcsa-monitoring/</guid><description>&lt;h2 id="vcenter-server-appliance-monitoring"&gt;vCenter Server Appliance Monitoring&lt;/h2&gt;&#10;&lt;h3 id="what-is-vcenter-server-appliance"&gt;What Is vCenter Server Appliance?&lt;/h3&gt;&#10;&lt;p&gt;vCenter Server Appliance (vCSA) is a powerful, preconfigured Linux-based virtual machine optimized for running VMware vCenter Server and associated services. It is a vital component in managing virtualized environments, providing centralized management of virtualized hosts and virtual machines from a single console.&lt;/p&gt;&#10;&lt;h3 id="monitoring-vcenter-server-appliance-with-netdata"&gt;Monitoring vCenter Server Appliance With Netdata&lt;/h3&gt;&#10;&lt;p&gt;Netdata offers a seamless and efficient way to monitor vCenter Server Appliance. As a comprehensive monitoring tool, it provides real-time insights and detailed metrics that are crucial for maintaining the health and performance of your vCSA environment. With Netdata’s &lt;a href="https://www.netdata.cloud/"&gt;free and open-source platform&lt;/a&gt;, you get simple configurations, interactive visualizations, and a low-overhead monitoring solution.&lt;/p&gt;</description></item><item><title>vCenter vpxd crash loop: the core service that keeps restarting</title><link>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vpxd-crash-loop/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-vpxd-crash-loop/</guid><description>&lt;p&gt;vpxd is the C++ core of vCenter Server. It holds the entire managed inventory in memory, dispatches every management task to ESXi hosts via hostd, runs DRS, executes statistics rollups, and serves every SDK client (vSphere Client, PowerCLI, Veeam, NSX Manager, Aria Operations, custom automation). When vpxd dies, vCenter is functionally down: no provisioning, no vMotion orchestration, no DRS, no HA reconfiguration. VMs already running on hosts keep running, and FDM still restarts them after a host failure, because HA does not depend on vpxd.&lt;/p&gt;</description></item><item><title>VMware vSphere Operations Guides</title><link>https://www.netdata.cloud/guides/vmware-vsphere/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/guides/vmware-vsphere/</guid><description>How vSphere and vCenter actually fail in production, what the early warning signals look like, and the runbooks for the symptoms you&amp;rsquo;ll see in real incidents — across both the ESXi data plane and the vCenter management plane.</description></item><item><title>vSphere monitoring checklist: the signals every host, VM, and vCenter needs</title><link>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-monitoring-checklist/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/guides/vmware-vsphere/vmware-vsphere-monitoring-checklist/</guid><description>&lt;p&gt;Send this to someone standing up vSphere monitoring for the first time, or rebuilding an alerting setup that pages too often and misses real incidents. It lists the signals worth collecting across the hypervisor plane (ESXi hosts and VMs) and the management plane (vCenter Server Appliance).&lt;/p&gt;&#10;&lt;p&gt;vSphere does not fail like a generic Linux box. CPU contention is invisible from inside the guest. Memory goes from fine to catastrophic in minutes once host swapping starts. A datastore at 99% full looks identical to one at 5% full from inside a VM, until every VM on it halts. And vCenter can degrade for weeks before anyone notices, because DRS, HA, and the API quietly keep working until they don&amp;rsquo;t. Generic CPU/disk/network dashboards miss most of this.&lt;/p&gt;</description></item></channel></rss>