<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Infrastructure-Monitoring on Netdata</title><link>https://www.netdata.cloud/tags/infrastructure-monitoring/</link><description>Recent content in Infrastructure-Monitoring on Netdata</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Mon, 29 Jun 2026 00:00:00 +0000</lastBuildDate><atom:link href="https://www.netdata.cloud/tags/infrastructure-monitoring/index.xml" rel="self" type="application/rss+xml"/><item><title>SolarWinds Observability Alternative: Real-Time &amp; AI</title><link>https://www.netdata.cloud/solarwinds-observability-alternative/</link><pubDate>Mon, 29 Jun 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solarwinds-observability-alternative/</guid><description>Netdata is the modern alternative to SolarWinds Hybrid Cloud Observability: per-second resolution, edge-native ML on every metric, zero-config deployment, and flat per-node pricing with full data sovereignty.</description></item><item><title>SolarWinds Orion Alternative: One Agent, Full Stack</title><link>https://www.netdata.cloud/solarwinds-orion-alternative/</link><pubDate>Mon, 29 Jun 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solarwinds-orion-alternative/</guid><description>Replace SolarWinds Orion Platform with Netdata: one agent, per-second visibility, ML anomaly detection, AI troubleshooting, and network flow analysis without stacked module licenses.</description></item><item><title>5 Best SolarWinds Alternatives for 2026</title><link>https://www.netdata.cloud/blog/solarwinds-alternatives-2026/</link><pubDate>Tue, 23 Jun 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/solarwinds-alternatives-2026/</guid><description>&lt;p>As organizations modernize their infrastructure and embrace cloud-native architectures, traditional monitoring solutions are showing their age. SolarWinds, while a long-established player in IT management, was designed for an era of static, on-premise infrastructure. In 2026, teams are seeking alternatives that can keep pace with dynamic, distributed systems—and they&amp;rsquo;re finding better options.&lt;/p>
&lt;h2 id="what-is-solarwinds-is-it-still-the-right-choice">What Is SolarWinds? Is It Still The Right Choice?&lt;/h2>
&lt;p>SolarWinds Platform (formerly Orion) has been a cornerstone of IT monitoring for decades, offering comprehensive coverage of network devices, servers, applications, and databases. It&amp;rsquo;s particularly strong in traditional enterprise environments with extensive hardware monitoring needs.&lt;/p></description></item><item><title>SolarWinds Price Increases 2026: What Customers Need to Know</title><link>https://www.netdata.cloud/blog/solarwinds-price-increases-2026/</link><pubDate>Tue, 23 Jun 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/solarwinds-price-increases-2026/</guid><description>&lt;p>If you&amp;rsquo;re a SolarWinds customer facing renewal, you&amp;rsquo;ve likely noticed significant changes to pricing and licensing terms in 2024-2025. You&amp;rsquo;re not alone. At Netdata, we&amp;rsquo;ve been speaking with dozens of SolarWinds customers who are reassessing their monitoring strategies in light of these changes. This post provides factual information about what&amp;rsquo;s changed, the real impact on organizations, and a practical framework for evaluating your path forward.&lt;/p>
&lt;h2 id="whats-happening-with-solarwinds-pricing">What&amp;rsquo;s Happening with SolarWinds Pricing?&lt;/h2>
&lt;p>In February 2025, SolarWinds was acquired by private equity firm Turn/River Capital in a $4.4 billion transaction. As is common with PE-backed acquisitions, this has led to significant changes in pricing and business terms. Based on customer reports and public information, renewal prices have increased by 100-300% for many customers. One customer on the SolarWinds community forum reported their renewal more than doubled, a 225% increase from the previous year.&lt;/p></description></item><item><title>The Best Infrastructure Monitoring Tools In 2026</title><link>https://www.netdata.cloud/resources/best-infrastructure-monitoring-tools/</link><pubDate>Sat, 30 May 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/resources/best-infrastructure-monitoring-tools/</guid><description/></item><item><title>Introducing the Netdata Cloud MCP Server</title><link>https://www.netdata.cloud/blog/netdata-cloud-mcp-server/</link><pubDate>Fri, 27 Feb 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-cloud-mcp-server/</guid><description>&lt;p>The Netdata Cloud MCP Server is now available — giving AI agents and assistants direct access to your Netdata through a single endpoint at &lt;code>app.netdata.cloud/api/v1/mcp&lt;/code>.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="ai-is-changing-how-we-monitor-infrastructure">AI Is Changing How We Monitor Infrastructure&lt;/h2>
&lt;p>If you&amp;rsquo;re an engineer in 2026, chances are AI is already part of your daily workflow, whether that&amp;rsquo;s a general-purpose assistant like ChatGPT, Claude, or Gemini that you bounce questions off, or a coding agent like &lt;a href="https://docs.anthropic.com/en/docs/claude-code">Claude Code&lt;/a>, &lt;a href="https://openai.com/index/codex/">Codex&lt;/a>, &lt;a href="https://www.cursor.com/">Cursor&lt;/a>, or &lt;a href="https://windsurf.com/">Windsurf&lt;/a> that writes and debugs code alongside you. These tools are incredibly powerful, but until now, they&amp;rsquo;ve been blind to what&amp;rsquo;s actually happening on your infrastructure.&lt;/p></description></item><item><title>Netdata vs Catchpoint | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/catchpoint/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/catchpoint/</guid><description>Netdata delivers real-time infrastructure visibility with per-second metrics, ML-based anomaly detection, and AI-powered troubleshooting - the essential internal monitoring layer that external testing platforms cannot provide.</description></item><item><title>Netdata vs Datadog | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/datadog/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/datadog/</guid><description/></item><item><title>Netdata vs Dynatrace | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/dynatrace/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/dynatrace/</guid><description>Netdata provides superior infrastructure monitoring at a fraction of Dynatrace&amp;rsquo;s cost, with true per-second granularity and complete data sovereignty. Learn when to use each platform and how a hybrid approach delivers complete observability at significantly lower total cost.</description></item><item><title>Netdata vs ELK Stack | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/elk/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/elk/</guid><description>A comprehensive comparison between Netdata and the ELK Stack for infrastructure monitoring and observability. Learn which solution fits your real-time monitoring, log management, and troubleshooting needs.</description></item><item><title>Netdata vs IBM Instana | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/instana/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/instana/</guid><description/></item><item><title>Netdata vs Last9 | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/last9/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/last9/</guid><description/></item><item><title>Netdata vs N-able: Real-Time Monitoring Comparison</title><link>https://www.netdata.cloud/comparisons/nable/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/nable/</guid><description>Netdata provides per-second infrastructure monitoring with ML-based anomaly detection and AI troubleshooting - capabilities N-able&amp;rsquo;s 5-10 minute intervals and basic dashboards can&amp;rsquo;t match. See how Netdata solves the monitoring gaps N-able customers experience daily.</description></item><item><title>Netdata vs New Relic | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/newrelic/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/newrelic/</guid><description/></item><item><title>Netdata vs Observium | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/observium/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/observium/</guid><description>Netdata vs Observium comparison: Real-time full-stack observability with ML and AI versus SNMP-based network device monitoring. See how Netdata&amp;rsquo;s distributed architecture, per-second granularity, and automated intelligence deliver comprehensive infrastructure visibility that network-only monitoring cannot provide.</description></item><item><title>Netdata vs Sumo Logic | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/sumologic/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/sumologic/</guid><description>Comprehensive comparison of Netdata and Sumo Logic: real-time monitoring, pricing models, deployment complexity, and use cases. Learn which platform fits your DevOps and observability needs.</description></item><item><title>Netdata vs UptimeRobot | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/uptimerobot/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/uptimerobot/</guid><description>Netdata and UptimeRobot serve different purposes: UptimeRobot monitors external availability while Netdata provides deep infrastructure diagnostics. Learn when to use each tool and how Netdata&amp;rsquo;s component-level alerts and ML-based anomaly detection provide deeper visibility.</description></item><item><title>Netdata vs WhatsUp Gold | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/whatsup-gold/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/whatsup-gold/</guid><description>Netdata vs WhatsUp Gold: Real-time observability with per-second granularity, zero configuration, and built-in ML/AI vs traditional Windows-centric network monitoring. See how Netdata delivers 90% cost reduction and 80% faster MTTR for modern infrastructure.</description></item><item><title>Netdata vs Zabbix | Monitoring Tools Comparison</title><link>https://www.netdata.cloud/comparisons/zabbix/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/comparisons/zabbix/</guid><description/></item><item><title>Proxmox Monitoring With Real-Time Observability</title><link>https://www.netdata.cloud/solutions/technologies/proxmox-monitoring/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solutions/technologies/proxmox-monitoring/</guid><description>Netdata delivers enterprise-grade Proxmox monitoring with per-second granularity, automated dashboards, and ML anomaly detection - all without the complexity or cost of traditional solutions.</description></item><item><title>Real-Time Infrastructure Monitoring For Freelancers</title><link>https://www.netdata.cloud/solutions/built-for/freelancers/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solutions/built-for/freelancers/</guid><description>Real-time infrastructure monitoring built for technical freelancers managing multiple client environments. Zero configuration, AI-powered insights, and predictable costs let you focus on delivering value, not managing monitoring tools.</description></item><item><title>Real-Time Observability For Developers Who Ship Fast</title><link>https://www.netdata.cloud/solutions/built-for/developers/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solutions/built-for/developers/</guid><description>Stop context switching between tools. Netdata provides developers with real-time infrastructure visibility, AI-assisted debugging, and zero-configuration monitoring—all in one unified platform that integrates directly into your IDE and workflow.</description></item><item><title>Real-Time Observability For Platform Engineers</title><link>https://www.netdata.cloud/solutions/built-for/platform-engineers/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solutions/built-for/platform-engineers/</guid><description>Empower platform engineering teams with Netdata&amp;rsquo;s distributed observability platform. Get per-second visibility, automated dashboards, ML anomaly detection, and predictable costs - all without query languages or complex pipelines.</description></item><item><title>Real-Time Observability For SRE Teams</title><link>https://www.netdata.cloud/solutions/built-for/sre/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solutions/built-for/sre/</guid><description>Transform SRE operations with Netdata&amp;rsquo;s edge-native observability platform. Get per-second visibility, ML anomaly detection on every metric, and AI-powered troubleshooting at 90% lower cost than traditional solutions.</description></item><item><title>Telecom Network Monitoring Software At Scale</title><link>https://www.netdata.cloud/solutions/industries/telecom/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solutions/industries/telecom/</guid><description>Transform telecom operations with distributed edge monitoring that delivers complete infrastructure visibility, ML-powered anomaly detection, and zero-configuration deployment in minutes.</description></item><item><title>VMware Monitoring Software Without Centralized Data</title><link>https://www.netdata.cloud/solutions/technologies/vmware-monitoring/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solutions/technologies/vmware-monitoring/</guid><description>Netdata delivers per-second VMware monitoring with zero configuration, built-in ML anomaly detection, and 90% cost savings. Monitor vSphere, ESXi hosts, VMs, and hybrid clouds with true real-time visibility.</description></item><item><title>Windows Server Monitoring Software At Scale</title><link>https://www.netdata.cloud/solutions/technologies/windows-monitoring/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/solutions/technologies/windows-monitoring/</guid><description>Transform Windows Server monitoring with Netdata&amp;rsquo;s edge-native platform. Get per-second visibility, automatic discovery, and AI-powered troubleshooting at a fraction of traditional costs.</description></item><item><title>AI Troubleshooting GA With On-Demand Credits</title><link>https://www.netdata.cloud/blog/ai-credits/</link><pubDate>Tue, 02 Sep 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/ai-credits/</guid><description>&lt;p>Since launching our AI investigations and insights in a research preview, one thing has become clear: &lt;strong>automated root cause analysis delivers a significant return on investment.&lt;/strong> Teams have confirmed that instant insights don&amp;rsquo;t just save a few minutes; they fundamentally shorten incident response cycles, free up valuable engineering hours, and reduce the business impact of downtime.&lt;/p>
&lt;p>The preview successfully demonstrated this value, with 10 free AI sessions per month allowing teams to integrate AI into their workflows. Now, based on the success and maturity of the capabilities, we are proud to announce that &lt;strong>Netdata&amp;rsquo;s AI investigations and insights are graduating from research preview to General Availability.&lt;/strong>&lt;/p></description></item><item><title>Netdata Now Troubleshoots Your Alerts for You</title><link>https://www.netdata.cloud/blog/automated-alert-troubleshooting/</link><pubDate>Sun, 03 Aug 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/automated-alert-troubleshooting/</guid><description>&lt;p>The 2 AM pager alert. For anyone in Ops, SRE, or IT administration, those words trigger a familiar sense of dread. An alert has fired. Is it a real fire, or another false alarm waking you from a dead sleep? The pressure is on. Every minute of downtime costs money and reputation, but troubleshooting a complex system when you&amp;rsquo;re sleep-deprived is a Herculean task.&lt;/p>
&lt;!--truncate-->
&lt;p>This cycle is a massive drain on engineering resources. The daily grind of sifting through alerts, trying to distinguish signal from noise, and manually correlating metrics to find a root cause consumes countless hours. This constant firefighting leads to alert fatigue, where even critical notifications start to get ignored. The core questions are always the same: Is this a real problem? What is the potential impact? Why did this trigger? What do I do next? Answering them is a slow, manual, and often stressful process.&lt;/p></description></item><item><title>Netdata vs Prometheus: A 2025 Performance Analysis</title><link>https://www.netdata.cloud/blog/netdata-vs-prometheus-2025/</link><pubDate>Thu, 23 Jan 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-vs-prometheus-2025/</guid><description>&lt;p>When it comes to infrastructure monitoring, performance, scalability, and efficiency are critical considerations. In this blog post, we revisit two widely adopted open-source monitoring solutions: &lt;strong>Netdata&lt;/strong> and &lt;strong>Prometheus&lt;/strong>. Both tools have introduced notable improvements in their latest versions, emphasizing scalability and enhanced efficiency.&lt;/p>
&lt;p>In our previous &lt;a href="https://www.netdata.cloud/blog/netdata-vs-prometheus-performance-analysis/">analysis&lt;/a>, we explored key differences between these systems, focusing on resource consumption and data retention. This follow-up expands on that comparison by subjecting both tools to a significantly larger workload. With the number of monitored nodes increased to 1000, containers to 80k, and metrics ingestion reaching 4.6 million metrics per second, we examine how each system performs under these demanding conditions, focusing on CPU utilization, memory requirements, disk I/O, network usage, and data retention during data ingestion.&lt;/p></description></item><item><title>5 Best Datadog Alternatives For Monitoring &amp; Observability</title><link>https://www.netdata.cloud/blog/5-datadog-alternatives/</link><pubDate>Tue, 15 Oct 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/5-datadog-alternatives/</guid><description>&lt;p>As businesses rely more on &lt;a href="https://www.netdata.cloud/academy/what-is-infrastructure-monitoring-and-why-you-need-it/">infrastructure monitoring&lt;/a>, observability tools have become essential for keeping systems secure, responsive, and efficient. In this evolving monitoring landscape, Datadog is known as a leading analytics and monitoring tool, capable of collecting essential performance metrics from servers, databases, applications, and other IT infrastructures. But at what cost?&lt;/p>
&lt;h2 id="what-is-datadog-is-it-your-only-option">What Is Datadog? Is It Your Only Option?&lt;/h2>
&lt;p>Datadog is a cloud-based monitoring platform that helps organizations track the performance of their infrastructure, applications, and logs. It provides visibility across systems to ensure efficient operations and identify issues quickly.&lt;/p></description></item><item><title>Exploring systemd journal logs with Netdata</title><link>https://www.netdata.cloud/blog/exploring-systemd-journal-logs/</link><pubDate>Thu, 12 Oct 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/exploring-systemd-journal-logs/</guid><description>&lt;p>Today, we released our &lt;code>systemd&lt;/code> &lt;strong>journal plugin for Netdata&lt;/strong>, allowing you to explore, view, search, filter and analyze &lt;code>systemd&lt;/code> journal logs.&lt;/p>
&lt;p>Like most things about Netdata, this is a &lt;strong>zero-configuration plugin&lt;/strong>. You don’t have to do anything apart from &lt;strong>installing Netdata&lt;/strong> on your systems.This is key design direction for Netdata, since we want Netdata to be able to help even if you install it mid-crisis, while you have an incident at hand.&lt;/p></description></item><item><title>systemd journal logs</title><link>https://www.netdata.cloud/blog/systemd-journal-logs/</link><pubDate>Mon, 09 Oct 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/systemd-journal-logs/</guid><description>&lt;p>&lt;em>“Why bother with it? I let it run in the background and focus on more important DevOps work.”&lt;/em>
— a random DevOps Engineer at Reddit r/devops&lt;/p>
&lt;p>In an era where technology is evolving at breakneck speeds, it&amp;rsquo;s easy to overlook the tools that are right under our noses. One such underutilized powerhouse is the &lt;strong>&lt;code>systemd&lt;/code> journal&lt;/strong>. For many, it&amp;rsquo;s a mere tool to check the status of systemd service units or to tail the most recent events (journalctl -f). Others who do mainly container work, ignore even its existence.&lt;/p></description></item><item><title>Netdata Cloud On Prem</title><link>https://www.netdata.cloud/blog/netdata-cloud-on-prem/</link><pubDate>Tue, 26 Sep 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-cloud-on-prem/</guid><description>&lt;p>We at &lt;strong>Netdata&lt;/strong> understand that &lt;a href="https://blog.netdata.cloud/future-of-infrastructure-monitoring/">infrastructure monitoring&lt;/a> can be a complex maze—high costs, specialized skill sets, scalability, data silos, and more. That&amp;rsquo;s why we have always aimed to streamline and modernize this critical operation. Today, we&amp;rsquo;re thrilled to announce the launch of &lt;a href="https://www.netdata.cloud/contact-us/?subject=on-prem">Netdata Cloud On Prem&lt;/a>, a ground-breaking solution designed for robust &lt;strong>on-prem infrastructure monitoring&lt;/strong> - it comes with all the &lt;strong>Netdata Cloud&lt;/strong> features you love but fully on prem.&lt;/p>
&lt;h2 id="netdata-cloud-on-prem">&lt;strong>Netdata Cloud On-Prem&lt;/strong>&lt;/h2>
&lt;p>While &lt;a href="https://www.netdata.cloud/">Netdata Cloud&lt;/a> never stores any of your metric data on the cloud and just streams it ephemerally while you view a chart, the demand for on premise infrastructure monitoring has never been more pressing. Many large enterprises, governmental organizations, research institutes and critical infrastructures require a level of data privacy, security, and customization that only an on-prem solution can offer.&lt;/p></description></item><item><title>The Future of Monitoring is Automated and Opinionated</title><link>https://www.netdata.cloud/blog/the-future-of-monitoring-is-automated-and-opinionated/</link><pubDate>Tue, 09 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/the-future-of-monitoring-is-automated-and-opinionated/</guid><description>&lt;p>So, you think you monitor your infra?&lt;/p>
&lt;!-- truncate -->
&lt;p>As humanity increasingly relies on technology, &lt;a href="https://www.netdata.cloud/blog/future-of-infrastructure-monitoring/">the need for reliable and efficient infrastructure monitoring solutions has never been greater&lt;/a>.&lt;/p>
&lt;p>However, most businesses don&amp;rsquo;t take this seriously. They make poor choices that soon trap their best talent, the people who should be propelling them ahead of their competition.&lt;/p>
&lt;p>Consider this: most of the world believes that each company needs to dedicate time, talent, and money to configure and set up the monitoring of their web servers and database servers from scratch!&lt;/p></description></item><item><title>Why Scalable Monitoring Matters For Modern Systems</title><link>https://www.netdata.cloud/blog/why-scalable-monitoring-is-essential-for-modern-distributed-systems/</link><pubDate>Wed, 26 Apr 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/why-scalable-monitoring-is-essential-for-modern-distributed-systems/</guid><description>&lt;p>&lt;img src="../2023-04-26-why-scalable-monitoring-is-essential/img/stacked-netdata.png" alt="stacked-netdata">&lt;/p>
&lt;p>It&amp;rsquo;s becoming increasingly common to discuss the importance of scalability in monitoring solutions and how it can impact the performance and reliability of distributed systems.&lt;/p>
&lt;!-- truncate -->
&lt;p>In today&amp;rsquo;s rapidly evolving technological landscape, organizations are increasingly relying on distributed systems to power their operations. These systems consist of multiple interconnected components that work together to deliver a cohesive experience. They can span across different geographic locations, and often involve a combination of &lt;a href="https://www.netdata.cloud/product/cloud-on-premises/">on-premises, cloud&lt;/a>, and &lt;a href="https://www.netdata.cloud/solutions/technologies/docker-monitoring/">container-based environments&lt;/a>. As such, effectively managing these complex systems is critical to ensuring optimal performance, reliability, and security.&lt;/p></description></item><item><title>Remote UNIX System Monitoring Using Net-SNMP</title><link>https://www.netdata.cloud/blog/remote-unix-monitoring-with-net-snmp/</link><pubDate>Wed, 12 Apr 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/remote-unix-monitoring-with-net-snmp/</guid><description>&lt;p>&lt;img src="../2023-04-12-remote-unix-monitoring-with-net-snmp/img/img.jpg" alt="img">&lt;/p>
&lt;p>Need to monitor a UNIX-like system, but can’t install Netdata on it? With our SNMP collector and Net-SNMP,
you can get basic system information with just a bit of relatively quick and easy configuration.&lt;/p>
&lt;!-- truncate -->
&lt;h2 id="what-is-snmp">What is SNMP?&lt;/h2>
&lt;p>The Simple Network Management Protocol, commonly known as SNMP, is a relatively lightweight protocol designed for
monitoring and configuration management for network appliances like switches, routers or gateways. However, it can also
be used for those purposes on almost any UNIX-like system thanks to the &lt;a href="http://www.net-snmp.org/">Net-SNMP project&lt;/a>.&lt;/p></description></item><item><title>Introducing the Netdata demo space</title><link>https://www.netdata.cloud/blog/netdata-demo/</link><pubDate>Fri, 24 Mar 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-demo/</guid><description>&lt;p>&lt;img src="https://user-images.githubusercontent.com/24860547/201481889-0cf8e192-683f-4a80-9b96-4f69dd85490f.png" alt="image">&lt;/p>
&lt;p>Introducing Netdata&amp;rsquo;s Demo Space, a quick and easy way to experience monitoring environments before you set them up yourself.&lt;/p>
&lt;!--truncate-->
&lt;p>At Netdata, we are always striving to provide the best monitoring experience for our users. We understand that adopting a new monitoring solution can sometimes be challenging, especially when you&amp;rsquo;re unsure of how it will fit your specific environment. That&amp;rsquo;s why we&amp;rsquo;re excited to announce the Netdata Demo Space!&lt;/p></description></item><item><title>How to monitor node reboots?</title><link>https://www.netdata.cloud/blog/monitoring-node-reboots/</link><pubDate>Thu, 17 Nov 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/monitoring-node-reboots/</guid><description>&lt;p>Monitoring the health and status of nodes and servers is a critical part of effective infrastructure monitoring.&lt;/p>
&lt;p>&lt;img src="https://user-images.githubusercontent.com/96257330/202475049-22838a0b-73b1-485b-8416-5fd49d6ccb53.png" alt="logo">&lt;/p>
&lt;!--truncate-->
&lt;h2 id="how-to-monitor-node-reboots">How to monitor node reboots?&lt;/h2>
&lt;p>One of the most critical tasks of monitoring an infrastructure is to check the health of its servers/nodes. In most cases, this results in setting up a &amp;ldquo;Hardware manager&amp;rdquo; from the hardware vendor delivering these servers or setting up an SNMP (or similar) agent to continuously monitor the availability of the server and report when there is a reboot / failure.&lt;/p></description></item><item><title>Monitoring &amp; troubleshooting Cassandra with Netdata</title><link>https://www.netdata.cloud/blog/cassandra-monitoring-part2/</link><pubDate>Sat, 29 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/cassandra-monitoring-part2/</guid><description>&lt;p>How to monitor and troubleshoot Cassandra with Netdata.&lt;/p>
&lt;p>&lt;img src="https://user-images.githubusercontent.com/24860547/198524087-37dda416-a9a9-4c55-b379-0f46e990f83f.png" alt="logo">&lt;/p>
&lt;!--truncate-->
&lt;p>&lt;em>&lt;strong>Note&lt;/strong>: This post is the second part of a Cassandra monitoring series. Be sure to read our first entry &lt;a href="https://blog.netdata.cloud/cassandra-monitoring-part1">here&lt;/a>.&lt;/em>&lt;/p>
&lt;h2 id="monitoring-cassandra-with-netdata">Monitoring Cassandra with Netdata&lt;/h2>
&lt;p>Netdata’s &lt;a href="https://learn.netdata.cloud/docs/agent/collectors/go.d.plugin/modules/cassandra">Cassandra collector documentation&lt;/a> explains how to set it up to collect metrics automatically.&lt;/p>
&lt;p>Once you have followed the instructions in the docs and have installed and configured Netdata on the Cassandra cluster you are ready to start monitoring and troubleshooting. Check out the &lt;a href="https://app.netdata.cloud/spaces/netdata-demo/rooms/cassandra/overview">Cassandra demo room&lt;/a> to interact with the charts, metrics and other functionality described here.&lt;/p></description></item><item><title>How to monitor and fix Database bloats in PostgreSQL?</title><link>https://www.netdata.cloud/blog/postgresql-database-bloat/</link><pubDate>Fri, 28 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/postgresql-database-bloat/</guid><description>&lt;p>Database bloat is disk space that was used by a table or index and is available for reuse by the database but has not been reclaimed. Bloat is created when deleting or updating tables and indexes. Here&amp;rsquo;s how to deal with it!&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-is-database-bloat">What is Database bloat?&lt;/h2>
&lt;p>Database bloat is disk space that was used by a table or index and is available for reuse by the database but has not been reclaimed. Bloat is created when deleting or updating tables and indexes.&lt;/p></description></item><item><title>Cassandra Monitoring: Key Metrics &amp; Best Practices</title><link>https://www.netdata.cloud/blog/cassandra-monitoring-part1/</link><pubDate>Thu, 27 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/cassandra-monitoring-part1/</guid><description>&lt;p>What are the important Cassandra metrics to monitor and how to monitor them.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-is-cassandra--why-use-it">What Is Cassandra &amp;amp; Why Use It&lt;/h2>
&lt;p>Cassandra is an open-source, distributed, wide-column NoSQL database management system written in Java. Cassandra was originally developed by &lt;a href="https://twitter.com/hedvigeng">Avinash Lakshmanan&lt;/a> and &lt;a href="https://twitter.com/pmalik">Prashant Malik&lt;/a> at Facebook and then released as open source, eventually becoming part of the &lt;a href="https://www.netdata.cloud/apache-monitoring/">Apache&lt;/a> project.&lt;/p>
&lt;p>&lt;a href="https://www.netdata.cloud/integrations/data-collection/databases/cassandra/">Cassandra is a NoSQL database&lt;/a> - NoSQL (also known as &amp;ldquo;not only SQL&amp;rdquo;) databases do not require data to be stored in tabular format. They provide flexible schemas and scale easily with large amounts of data and high user loads.&lt;/p></description></item><item><title>How to find out which application is causing server load</title><link>https://www.netdata.cloud/blog/server-load/</link><pubDate>Wed, 26 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/server-load/</guid><description>&lt;p>We often hear the term load used to describe the state of a server or a device, but we&amp;rsquo;re here to tell you what it means, precisely, and how to monitor it.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-is-server-load">What is server load?&lt;/h2>
&lt;p>We often hear the term &amp;ldquo;load&amp;rdquo; used to describe the state of a server or a &lt;a href="https://www.netdata.cloud/blog/iot-monitoring-challenges/">device&lt;/a>. But what does it really mean?&lt;/p>
&lt;p>System load is a measure of the amount of computational work that a system performs. An overloaded system, by definition, isn&amp;rsquo;t able to complete all its
tasks per schedule - this affects the performance and productivity of the system. And while &amp;ldquo;load&amp;rdquo; often gets conflated with CPU usage there&amp;rsquo;s a lot more to it.&lt;/p></description></item><item><title>How to monitor the disk usage on your infrastructure</title><link>https://www.netdata.cloud/blog/disk-usage/</link><pubDate>Tue, 25 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/disk-usage/</guid><description>&lt;p>The most important part of disk usage monitoring is to check the utilization of each filesystem and each mount point which can reveal existing or impending issues with the storage space on your infrastructure.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-does-disk-usage-du-mean">What Does Disk Usage (DU) Mean?&lt;/h2>
&lt;p>Disk usage (DU) refers to the portion or percentage of computer storage that is currently in use. It contrasts with disk space or &lt;a href="https://www.netdata.cloud/blog/server-uptime-monitoring-why-do-we-need-it/">capacity&lt;/a>,
which is the total amount of space that a given disk is capable of storing. Disk usage is a crucial metric to any computing system,
as it gives the user the information needed not only for storage, but also software requirements and overall operation. Although it usually
refers to a computer’s hard disk, it may also refer to external storage, such as a USB drive or compact disc (CD).&lt;/p></description></item><item><title>7 types of Redis latency and how to fix it</title><link>https://www.netdata.cloud/blog/7-types-of-redis-latency/</link><pubDate>Mon, 24 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/7-types-of-redis-latency/</guid><description>&lt;p>Redis is designed to be fast. In most cases, it is. However, there are times when Redis may be slow, due to network issues, disk latency, or other factors. When this happens, it is important to be able to &lt;a href="https://www.netdata.cloud/blog/server-uptime-monitoring-why-do-we-need-it/">detect the slow down&lt;/a> and investigate the cause of Redis latency.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="redis--latency">Redis &amp;amp; latency&lt;/h2>
&lt;p>Latency is the maximum delay between the time a client issues a command and the time the reply to the command is received by the client. Redis has strict requirements on average and worst case latency.&lt;/p></description></item><item><title>How to monitor systemd service liveness</title><link>https://www.netdata.cloud/blog/systemd-service-liveness/</link><pubDate>Fri, 21 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/systemd-service-liveness/</guid><description>&lt;p>The life of a sysadmin or SRE is often difficult, but occasionally very simple things can make a huge difference. Basic monitoring of your systemd services is one of those simple things, which we sometimes overlook. The simplest question one would want to know is if the thing that’s supposed to be running is actually running at all. If you use systemd services, you can guarantee an answer to that question within minutes using Netdata.&lt;/p></description></item><item><title>Web Servers Monitoring: Key Metrics &amp; Strategies</title><link>https://www.netdata.cloud/blog/web-servers-and-their-performance/</link><pubDate>Thu, 20 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/web-servers-and-their-performance/</guid><description>&lt;!--truncate-->
&lt;h2 id="the-importance-of-monitoring-web-servers">The Importance Of Monitoring Web Servers&lt;/h2>
&lt;p>Web servers are among the most important components in modern &lt;a href="https://www.netdata.cloud/blog/future-of-infrastructure-monitoring/">IT infrastructures&lt;/a>. They host the websites, web services, and &lt;a href="https://www.netdata.cloud/dotnet-monitoring/">web applications&lt;/a> that we use on a daily basis. Social networking, media streaming, software as a service (SaaS), and other activities wouldn’t be possible without the use of web servers. And with the advent of cloud computing and the movement of more services online, &lt;a href="https://www.netdata.cloud/blog/server-uptime-monitoring-why-do-we-need-it/">web servers and their monitoring are only becoming more important&lt;/a>. Given the extensive usage of Web servers, Sysadmins and SREs should monitor web servers as a key aspect for performance. &lt;/p></description></item><item><title>How to monitor HTTP endpoints</title><link>https://www.netdata.cloud/blog/http-endpoints/</link><pubDate>Mon, 17 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/http-endpoints/</guid><description>&lt;p>The &lt;a href="https://en.wikipedia.org/wiki/Hypertext_Transfer_Protocol">HTTP protocol&lt;/a> has become the de facto standard application layer protocol of the internet. From publicly available web sites and APIs to “inter-process” communications in REST based microservice architectures or large &lt;a href="https://en.wikipedia.org/wiki/Service-oriented_architecture">Service Oriented Architectures&lt;/a> based on &lt;a href="https://en.wikipedia.org/wiki/SOAP">SOAP&lt;/a>, you find HTTP being used again and again, due to its simplicity and our familiarity with it. How many protocols can you name that have &lt;a href="https://imgur.com/gallery/4KqWq">memes&lt;/a> for their status codes? Of course, such a popular protocol has endless pages written about how to properly monitor the services that rely on it, with many options specific to every use case.&lt;!--truncate--> What you will learn here is how to get your basics done in monitoring HTTP endpoints, so you can be up and running in a few minutes, monitoring all HTTP services in your &lt;a href="https://www.netdata.cloud/blog/future-of-infrastructure-monitoring/">entire infrastructure&lt;/a>. &lt;/p></description></item><item><title>How to monitor DNS query response time</title><link>https://www.netdata.cloud/blog/dns-query-response-time/</link><pubDate>Wed, 12 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/dns-query-response-time/</guid><description>&lt;p>DNS (Domain Name System) servers translate standard language web addresses to their actual IP addresses for network access.&lt;/p>
&lt;p>&lt;a href="https://xiaolishen.medium.com/the-dns-lookup-journey-240e9a5d345c">&lt;strong>DNS Lookup Journey&lt;/strong>&lt;/a>
&lt;img src="../wp-archive/uploads/2022/10/DNS-1.png" alt="">&lt;/p>
&lt;!--truncate-->
&lt;p>DNS response time is the time it takes a Domain Name Server to receive the request for a domain name’s IP address, process it, and return the IP address to the browser or application requesting it. When it comes to DNS response times, the lower the better, and generally values less than 100ms are considered to be in the acceptable range (depending on the application).&lt;/p></description></item><item><title>Why is data replication important?</title><link>https://www.netdata.cloud/blog/why-is-data-replication-important/</link><pubDate>Wed, 12 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/why-is-data-replication-important/</guid><description>&lt;p>High availability. This is what every monitoring tool needs to ensure that you never compromise on IT infrastructure visibility.&lt;!--truncate--> On top of high availability, do you really want to enable all available features on your production system? It is important for the monitoring tool to have a low footprint on your CPU consumption and memory usage. Let’s dive deeper into the recommended way of configuring Netdata to ensure high availability and a low resource footprint through data replication.&lt;/p></description></item><item><title>Data Collection Strategies For Infrastructure</title><link>https://www.netdata.cloud/blog/data-collection-strategies/</link><pubDate>Tue, 06 Sep 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/data-collection-strategies/</guid><description>&lt;!--truncate-->
&lt;p>Monitoring and troubleshooting; unfortunately, these terms are still used interchangeably, which can lead to misunderstandings about data collection strategies.&lt;/p>
&lt;p>In this article we aim to clarify some important definitions, processes, and common data collection strategies for monitoring solutions. We will specify the limitations of the described strategies, as well as key benefits which can potentially be also used for troubleshooting needs.&lt;/p>
&lt;p>&lt;strong>IT infrastructure monitoring&lt;/strong> is a business process of collecting and analyzing data over a period of time to improve business results.&lt;/p></description></item><item><title>Monitoring without Cooperation: Kubernetes</title><link>https://www.netdata.cloud/blog/monitoring-without-cooperation-kubernetes/</link><pubDate>Fri, 20 May 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/monitoring-without-cooperation-kubernetes/</guid><description>&lt;!--truncate-->
&lt;iframe width="560" height="315" src="https://www.youtube.com/embed/J2kdSTRJzV4?si=Rqv4gpqz_wZKusaZ" title="YouTube video player" frameborder="0" allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" referrerpolicy="strict-origin-when-cross-origin" allowfullscreen>&lt;/iframe>
&lt;p>Imagine this: You are an engineer at a startup. You are responsible for keeping all the applications running smoothly and safely in production. At first, you have things under control, but soon enough things start getting more complex. The system has grown by hundreds of new server nodes, new services and containers are showing up seemingly every day, and you hear there’s a project to transition everything to Kubernetes (you’ve heard the term “cloud native” so often that the phrase shows up in your dreams).&lt;/p></description></item><item><title>Kubernetes Throttling Doesn’t Have To Suck. Let Us Help!</title><link>https://www.netdata.cloud/blog/kubernetes-throttling-doesnt-have-to-suck-let-us-help/</link><pubDate>Tue, 03 May 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/kubernetes-throttling-doesnt-have-to-suck-let-us-help/</guid><description>&lt;p>CPU limits are probably the most misunderstood concept in Kubernetes CPU resources allocation and management.&lt;/p>
&lt;!--truncate-->
&lt;p>A lot of engineers advise the use of CPU limits on every container as Kubernetes best practice. Unfortunately, as we will prove below, they are wrong: CPU limits should rarely be used, if used at all!&lt;/p>
&lt;p>But why? What are the reasons that even senior DevOps engineers with vast experience in the field advise the use of CPU limits?&lt;/p></description></item><item><title>What is Infrastructure Monitoring?</title><link>https://www.netdata.cloud/blog/what-is-infrastructure-monitoring/</link><pubDate>Tue, 30 Jun 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/what-is-infrastructure-monitoring/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-medium wp-image-16343" src="../wp-archive/uploads/2022/03/Blog-What_is_Infrastructure_Monitoring_Header-600x450.png" alt="" width="600" height="450" />
&lt;p>IT is advancing blazingly fast. To keep up with architectural changes and hybrid environments, it’s more important than ever to maintain efficient infrastructure monitoring and troubleshooting. Adding to the complexity is the increase of distributed systems, comprised of many components and services. For IT teams to effectively manage monitoring modern infrastructure, it’s necessary to have the right practices and tools in place that enable teams to do their jobs as quickly as possible with fewer resources.&lt;/p></description></item></channel></rss>