<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>HPC on Netdata</title><link>https://www.netdata.cloud/tags/hpc/</link><description>Recent content in HPC on Netdata</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Sun, 23 Aug 2026 02:03:02 +0300</lastBuildDate><atom:link href="https://www.netdata.cloud/tags/hpc/index.xml" rel="self" type="application/rss+xml"/><item><title>HPC Monitoring Software With Per-Second Metrics</title><link>https://www.netdata.cloud/solutions/use-cases/hpc/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solutions/use-cases/hpc/</guid><description>Transform HPC operations with distributed edge-native monitoring that delivers sub-2-second insights, 90% cost reduction, and linear scalability to 100,000+ nodes - without the complexity.</description></item><item><title>Network Monitoring Software For Education</title><link>https://www.netdata.cloud/solutions/industries/education/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solutions/industries/education/</guid><description>Netdata delivers comprehensive observability for educational institutions with zero-configuration deployment, complete data sovereignty, and 90% cost savings. Monitor everything from learning management systems to research clusters with per-second precision.</description></item><item><title>What Is An HPC Cluster - Key Components &amp; How It Works</title><link>https://www.netdata.cloud/academy/what-is-an-hpc-cluster/</link><pubDate>Wed, 01 May 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/academy/what-is-an-hpc-cluster/</guid><description>&lt;p&gt;Imagine needing to analyze petabytes of genomic data, simulate the airflow over a new aircraft wing, or predict complex financial market movements. A single computer, no matter how powerful, would struggle or take an impractically long time. This is where High-Performance Computing (HPC) clusters come in. They provide the immense computational power needed to tackle problems far beyond the reach of standard computing. If you&amp;rsquo;re stepping into roles involving large-scale data processing or complex simulations, understanding HPC clusters is essential.&lt;/p&gt;</description></item><item><title>University of Calgary: Netdata in Academia</title><link>https://www.netdata.cloud/case-studies/education/calgary-university/</link><pubDate>Tue, 26 Mar 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/case-studies/education/calgary-university/</guid><description>&lt;h2 id="empowering-academic-research-with-real-time-monitoring"&gt;Empowering Academic Research with Real-Time Monitoring&lt;/h2&gt;&#10;&lt;p&gt;At the &lt;strong&gt;Machine Learning Lab&lt;/strong&gt; at the &lt;strong&gt;University of Calgary&lt;/strong&gt;, managing a robust infrastructure of servers and workstations is critical for advancing their research in machine learning (ML). Assistant Professor &lt;strong&gt;Yani Ioannou&lt;/strong&gt; oversees this infrastructure, ensuring that these essential resources remain operational and are used to their fullest potential. The lab&amp;rsquo;s challenge lies not just in the maintenance of these resources but in minimizing downtime to keep the research moving forward.&lt;/p&gt;</description></item><item><title>Enhancing Research Through Advanced Monitoring</title><link>https://www.netdata.cloud/case-studies/education/landcare/</link><pubDate>Thu, 01 Feb 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/case-studies/education/landcare/</guid><description>&lt;h2 id="helping-scientists-focus-on-the-science"&gt;Helping Scientists focus on the Science&lt;/h2&gt;&#10;&lt;p&gt;At &lt;a href="https://www.landcareresearch.co.nz/"&gt;Manaaki Whenua – Landcare Research&lt;/a&gt;, managing multi-user Linux workstations for scientific computing presented a unique challenge. Tasked with overseeing system utilization without a dedicated role in the organization, the need for a solution that was intuitive and efficient became paramount. Netdata&amp;rsquo;s simplicity in setup and the ability to monitor systems remotely via the cloud addressed these challenges head-on, offering a seamless solution for workstation management.&lt;/p&gt;</description></item><item><title>Lustre Metadata Monitoring</title><link>https://www.netdata.cloud/monitoring-101/lustre-monitoring/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/monitoring-101/lustre-monitoring/</guid><description>&lt;h2 id="lustre-metadata-monitoring"&gt;Lustre Metadata Monitoring&lt;/h2&gt;&#10;&lt;h3 id="what-is-lustre-metadata"&gt;What Is Lustre Metadata?&lt;/h3&gt;&#10;&lt;p&gt;Lustre is a type of parallel distributed file system, widely used for large-scale cluster computing. Originally developed for research and enterprise sectors due to its capacity and high performance, Lustre metadata refers to the internal management data that keeps track of file location, size, and storage attributes. Monitoring Lustre metadata is crucial for maintaining optimal file system operations and ensuring efficient management.&lt;/p&gt;&#10;&lt;h3 id="monitoring-lustre-metadata-with-netdata"&gt;Monitoring Lustre Metadata With Netdata&lt;/h3&gt;&#10;&lt;p&gt;Monitoring Lustre with Netdata provides seamless tracking of all critical metrics using an openmetrics (prometheus) exporter. To monitor Lustre metadata, Netdata taps into the &lt;a href="https://github.com/GSI-HPC/prometheus-cluster-exporter"&gt;Cluster Exporter&lt;/a&gt; which is capable of gathering essential data. With Netdata, you can ingest data from any Prometheus exporter; unlocking automated dashboards and alerts without needing a Prometheus server or Grafana. This integration supports efficient Lustre metadata monitoring and ensures constant observability over your cluster&amp;rsquo;s performance.&lt;/p&gt;</description></item><item><title>Slurm Monitoring</title><link>https://www.netdata.cloud/monitoring-101/slurm-monitoring/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/monitoring-101/slurm-monitoring/</guid><description>&lt;h2 id="slurm-monitoring"&gt;Slurm Monitoring&lt;/h2&gt;&#10;&lt;h3 id="what-is-slurm"&gt;What Is Slurm?&lt;/h3&gt;&#10;&lt;p&gt;Slurm, also known as the Simple Linux Utility for Resource Management, is an open-source workload management system that is specifically tailored for high-performance computing (HPC) and cluster environments. It efficiently allocates resources such as CPU and memory to various jobs, ensuring optimal use of available resources across clustered nodes.&lt;/p&gt;&#10;&lt;h3 id="monitoring-slurm-with-netdata"&gt;Monitoring Slurm With Netdata&lt;/h3&gt;&#10;&lt;p&gt;To effectively monitor Slurm, Netdata utilizes an openmetrics (Prometheus) exporter called the &lt;a href="https://github.com/vpenso/prometheus-slurm-exporter"&gt;Prometheus Slurm Exporter&lt;/a&gt;. With Netdata, you can ingest data from any Prometheus exporter, streamlining the process by providing automated dashboards, real-time alerts, and comprehensive insights without the need for setting up a standalone Prometheus server or configuring Grafana.&lt;/p&gt;</description></item></channel></rss>