<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Blogs on Netdata</title><link>https://www.netdata.cloud/blog/</link><description>Recent content in Blogs on Netdata</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Wed, 15 Jul 2026 00:00:00 +0000</lastBuildDate><atom:link href="https://www.netdata.cloud/blog/index.xml" rel="self" type="application/rss+xml"/><item><title>Native macOS Monitoring: Logs, Sensors, GPU &amp; Hardware Health</title><link>https://www.netdata.cloud/blog/macos-monitoring/</link><pubDate>Wed, 15 Jul 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/macos-monitoring/</guid><description>&lt;p>&lt;img src="../images/macos-monitoring.svg" alt="Native macOS monitoring with Netdata: unified logs, power, sensors, GPU, per-app metrics, storage, and network">&lt;/p>
&lt;p>We&amp;rsquo;ve overhauled macOS monitoring in the latest Netdata release. Netdata already collects system metrics on Macs at per-second resolution; this release completes the picture with logs and hardware telemetry, areas that previously required users to run CLI tools like &lt;code>log show&lt;/code> and &lt;code>powermetrics&lt;/code>. The new collectors read this data through Apple&amp;rsquo;s own frameworks, allowing users to trace application and OS errors and catch hardware issues early.&lt;/p></description></item><item><title>Fleet Observability: Linux Edge Device Monitoring</title><link>https://www.netdata.cloud/blog/fleet-observability/</link><pubDate>Sun, 28 Jun 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/fleet-observability/</guid><description>&lt;p>It feels less like managing devices and more like remote babysitting. You check the dashboard, everything is green, and then a customer in the field tells you a device has been down for two days. At a handful of servers, the rare failure is an event. Across thousands of distributed Linux endpoints — robots in warehouses, EV chargers across a city, kiosks in retail, IoT gateways in the field — the rare failure becomes a daily occurrence, and the tools built for a datacenter quietly stop telling you the truth.&lt;/p></description></item><item><title>Real Time Network Monitoring: Topology, NetFlow, SNMP</title><link>https://www.netdata.cloud/blog/network-monitoring/</link><pubDate>Wed, 24 Jun 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/network-monitoring/</guid><description>&lt;p>Interface counters tell you a port is busy. Bytes in, bytes out, errors, drops. That&amp;rsquo;s enough to know a link is saturated, but not enough to know which conversations are saturating it, which devices are involved, or how a problem propagates across your network. For that you&amp;rsquo;ve traditionally needed dedicated network performance monitoring tools, usually expensive, usually a separate console from the rest of your monitoring.&lt;/p>
&lt;p>Today we&amp;rsquo;re closing that gap. Netdata has had solid network interface monitoring for a long time through its native collectors and SNMP support. We&amp;rsquo;ve now built out the rest of the picture, and it adds up to NPM-class network monitoring: live network topology, NetFlow and sFlow traffic analysis, SNMP device monitoring across 200+ vendor profiles, SNMP trap handling, and a dedicated network monitoring dashboard. All of it runs alongside the infrastructure, application, and container metrics Netdata already collects, on the same timeline, in the same platform.&lt;/p></description></item><item><title>5 Best SolarWinds Alternatives for 2026</title><link>https://www.netdata.cloud/blog/solarwinds-alternatives-2026/</link><pubDate>Tue, 23 Jun 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/solarwinds-alternatives-2026/</guid><description>&lt;p>As organizations modernize their infrastructure and embrace cloud-native architectures, traditional monitoring solutions are showing their age. SolarWinds, while a long-established player in IT management, was designed for an era of static, on-premise infrastructure. In 2026, teams are seeking alternatives that can keep pace with dynamic, distributed systems—and they&amp;rsquo;re finding better options.&lt;/p>
&lt;h2 id="what-is-solarwinds-is-it-still-the-right-choice">What Is SolarWinds? Is It Still The Right Choice?&lt;/h2>
&lt;p>SolarWinds Platform (formerly Orion) has been a cornerstone of IT monitoring for decades, offering comprehensive coverage of network devices, servers, applications, and databases. It&amp;rsquo;s particularly strong in traditional enterprise environments with extensive hardware monitoring needs.&lt;/p></description></item><item><title>SolarWinds Price Increases 2026: What Customers Need to Know</title><link>https://www.netdata.cloud/blog/solarwinds-price-increases-2026/</link><pubDate>Tue, 23 Jun 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/solarwinds-price-increases-2026/</guid><description>&lt;p>If you&amp;rsquo;re a SolarWinds customer facing renewal, you&amp;rsquo;ve likely noticed significant changes to pricing and licensing terms in 2024-2025. You&amp;rsquo;re not alone. At Netdata, we&amp;rsquo;ve been speaking with dozens of SolarWinds customers who are reassessing their monitoring strategies in light of these changes. This post provides factual information about what&amp;rsquo;s changed, the real impact on organizations, and a practical framework for evaluating your path forward.&lt;/p>
&lt;h2 id="whats-happening-with-solarwinds-pricing">What&amp;rsquo;s Happening with SolarWinds Pricing?&lt;/h2>
&lt;p>In February 2025, SolarWinds was acquired by private equity firm Turn/River Capital in a $4.4 billion transaction. As is common with PE-backed acquisitions, this has led to significant changes in pricing and business terms. Based on customer reports and public information, renewal prices have increased by 100-300% for many customers. One customer on the SolarWinds community forum reported their renewal more than doubled, a 225% increase from the previous year.&lt;/p></description></item><item><title>High Cardinality Metrics At Scale: A Better Playbook</title><link>https://www.netdata.cloud/blog/high-cardinality-metrics-observability-scale/</link><pubDate>Sat, 30 May 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/high-cardinality-metrics-observability-scale/</guid><description>&lt;p>The &amp;ldquo;high cardinality is expensive&amp;rdquo; sentence has become observability&amp;rsquo;s version of &amp;ldquo;in this economy&amp;rdquo;: said so often that nobody questions whether it&amp;rsquo;s true. Every vendor pricing page invokes it. Every glossary article repeats it. Every architecture diagram shows aggregation buffers placed &lt;em>before&lt;/em> the storage layer.&lt;/p>
&lt;!--truncate-->
&lt;p>&amp;ldquo;High cardinality is expensive&amp;rdquo; is not a fact about the universe; it&amp;rsquo;s a fact about one architectural choice: centralizing time-series storage and querying it through an index that scales with unique series. Once you accept that choice, everything follows. You pay per metric, you drop labels you wish you&amp;rsquo;d kept, you pre-aggregate before storage, and you discover that the bug you were debugging only existed at the full resolution you already threw away.&lt;/p></description></item><item><title>Netdata Skills: Teach Your AI Coding Agent To Monitor</title><link>https://www.netdata.cloud/blog/netdata-skills/</link><pubDate>Wed, 20 May 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-skills/</guid><description>&lt;p>There&amp;rsquo;s a growing ecosystem of AI coding agents: Claude Code, Cursor, Copilot, Codex, Gemini CLI, Windsurf, and others. They&amp;rsquo;re good at writing code, but they don&amp;rsquo;t inherently know how to instrument that code for observability, configure monitoring infrastructure, or troubleshoot production systems using real telemetry data. That knowledge lives in documentation, runbooks, and the heads of your senior SREs.&lt;/p>
&lt;p>We&amp;rsquo;ve open-sourced a repository that encodes this knowledge into a format AI agents can use directly. &lt;a href="https://github.com/netdata/skills">netdata/skills&lt;/a> is a collection of agent skills, published in the open &lt;a href="https://agentskills.io">agentskills.io&lt;/a> format, that teach AI coding agents how to set up Netdata, instrument applications with OpenTelemetry, build collector pipelines, troubleshoot 49 specific technologies, and verify everything against live data via MCP.&lt;/p></description></item><item><title>OpenTelemetry and Netdata, Today</title><link>https://www.netdata.cloud/blog/opentelemetry-metrics-and-logs-ingestion/</link><pubDate>Fri, 15 May 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/opentelemetry-metrics-and-logs-ingestion/</guid><description>&lt;p>OpenTelemetry has become the default way to instrument applications and ship telemetry. The hard part has never been the data model. It&amp;rsquo;s been picking a backend that handles OTLP without quietly turning into a per-metric bill or a black box that swallows your data.&lt;/p>
&lt;p>Netdata is a native OTLP backend. Stand up an OpenTelemetry Collector with any of its hundreds of receivers, point its OTLP exporter at Netdata, and you get per-second charts, ML anomaly detection on every signal, AI-assisted troubleshooting, and infrastructure correlation, with no per-metric, per-series, or per-host charges. Metrics and logs work today. Trace support is coming soon.&lt;/p></description></item><item><title>Dashboard Playlists: Cycle Through Dashboards in TV Mode</title><link>https://www.netdata.cloud/blog/dashboard-playlists/</link><pubDate>Tue, 12 May 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/dashboard-playlists/</guid><description>&lt;p>When we shipped TV mode, we heard almost immediately: &amp;ldquo;Great, but I have five dashboards and one screen.&amp;rdquo; A single dashboard on a wall display covers one view of your infrastructure. If you want to rotate between your network overview, database health, application metrics, and infrastructure summary, someone has to walk over and click, or you&amp;rsquo;re buying more screens.&lt;/p>
&lt;p>Dashboard playlists solve this. You can now select a sequence of dashboards to cycle through in TV mode, with a configurable rotation interval. Set it up once, open the TV mode URL on your display, and the screen rotates through your chosen dashboards on its own.&lt;/p></description></item><item><title>Azure Local Migration: Monitor Both Sides In One View</title><link>https://www.netdata.cloud/blog/azure-local-migration/</link><pubDate>Mon, 11 May 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/azure-local-migration/</guid><description>&lt;p>&lt;img src="../images/azure-local-migration.svg" alt="Monitoring Your Azure to Azure Local Migration: One Dashboard for Both Sides">&lt;/p>
&lt;p>More organizations are moving workloads from Azure public cloud to Azure Local (formerly Azure Stack HCI) than most people realize. The reasons vary: data sovereignty requirements, latency-sensitive workloads that need to be closer to the edge, cost optimization for predictable workloads where reserved cloud capacity doesn&amp;rsquo;t make financial sense, or regulatory constraints that require data to stay on-premises.&lt;/p></description></item><item><title>Geo Maps: See Where Your Infrastructure Lives</title><link>https://www.netdata.cloud/blog/geo-maps/</link><pubDate>Sat, 09 May 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/geo-maps/</guid><description>&lt;p>When your infrastructure is spread across regions, data centers, branch offices, or edge locations, knowing where a node is physically located matters more than people usually admit. During an incident, &amp;ldquo;the node in the Singapore POP&amp;rdquo; communicates faster than a hostname. When you&amp;rsquo;re planning capacity, seeing geographic clustering tells you something that a flat list of nodes doesn&amp;rsquo;t. When a subset of your fleet starts misbehaving, the first question is often &amp;ldquo;is this regional?&amp;rdquo;&lt;/p></description></item><item><title>NVIDIA DCGM Collector: Deep GPU Monitoring For AI</title><link>https://www.netdata.cloud/blog/nvidia-dcgm-monitoring/</link><pubDate>Mon, 04 May 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/nvidia-dcgm-monitoring/</guid><description>&lt;p>&lt;img src="../images/dcgm-collector.svg" alt="NVIDIA DCGM Collector: Deep GPU Monitoring for Data Center and AI Infrastructure">&lt;/p>
&lt;p>GPU infrastructure is expensive and increasingly central to production workloads. Whether you&amp;rsquo;re running ML training jobs, inference serving, video transcoding, or HPC workloads, understanding what your GPUs are actually doing, and what&amp;rsquo;s going wrong when performance degrades, is not optional. The problem is that NVIDIA&amp;rsquo;s Data Center GPU Manager (DCGM) exposes an enormous amount of telemetry, but getting that data into a monitoring system in a useful, organized way has traditionally required significant setup and custom dashboarding work.&lt;/p></description></item><item><title>Misconfigured Alert Detection: Tuning Made Easy</title><link>https://www.netdata.cloud/blog/identifying-misconfigured-alerts/</link><pubDate>Tue, 28 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/identifying-misconfigured-alerts/</guid><description>&lt;p>&lt;img src="../images/miscon-alerts-1.png" alt="Misconfigured Alert Detection: Find the Alerts That Need Tuning">&lt;/p>
&lt;p>Netdata ships with hundreds of stock alerts. They cover a wide range of infrastructure conditions and they&amp;rsquo;re designed with sensible defaults. But &amp;ldquo;sensible defaults&amp;rdquo; and &amp;ldquo;correct for your environment&amp;rdquo; are not the same thing. A CPU threshold that&amp;rsquo;s perfectly reasonable for a build server might generate constant noise on a machine running batch jobs. An alert that&amp;rsquo;s critical for production might be irrelevant in staging, where it fires daily and everyone ignores it.&lt;/p></description></item><item><title>Azure Monitor Collector: Monitor Azure Infrastructure</title><link>https://www.netdata.cloud/blog/azure-monitor/</link><pubDate>Mon, 27 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/azure-monitor/</guid><description>&lt;p>&lt;img src="../images/azure-monitor-collector.svg" alt="Azure Monitor Collector: Monitor Your Entire Azure Infrastructure From Netdata">&lt;/p>
&lt;p>If you&amp;rsquo;re running infrastructure on Azure, you&amp;rsquo;ve probably dealt with the split between your Azure-native monitoring and the rest of your stack. Your VMs, databases, and Kubernetes clusters generate platform metrics through Azure Monitor, but those metrics live in a separate world from the OS-level, application, and on-prem metrics you&amp;rsquo;re already watching in Netdata. You end up checking two (or more) places during incidents, building mental bridges between dashboards that don&amp;rsquo;t talk to each other.&lt;/p></description></item><item><title>Database Performance Monitoring: 14+ DBs Supported</title><link>https://www.netdata.cloud/blog/dbm/</link><pubDate>Fri, 24 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/dbm/</guid><description>&lt;p>&lt;img src="../images/dbm-hero.svg" alt="Database Performance Monitoring: Query-Level Visibility Across 14+ Databases">&lt;/p>
&lt;p>Netdata has always collected database metrics: connections, throughput, replication lag, buffer cache hit ratios, and so on. These tell you that something is wrong, but they don&amp;rsquo;t tell you why. When your PostgreSQL response time spikes, the metric alone doesn&amp;rsquo;t tell you which query is responsible. For that, you&amp;rsquo;ve traditionally needed to SSH into the box, connect to the database, and run diagnostic queries manually. Or set up a separate database monitoring tool entirely.&lt;/p></description></item><item><title>Nagios Plugins: Run Existing Checks &amp; Custom Scripts</title><link>https://www.netdata.cloud/blog/nagios-plugins/</link><pubDate>Wed, 22 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/nagios-plugins/</guid><description>&lt;p>A lot of teams have a collection of Nagios plugins and custom monitoring scripts that have been running reliably for years. Some are standard community plugins for checking disk health or SSL certificate expiry. Others are homegrown Bash or Python scripts that check something very specific to the business: whether an API endpoint returns the right payload, whether a batch job completed on time, whether a queue depth is within bounds. These scripts work, they&amp;rsquo;re battle-tested, and nobody wants to rewrite them.&lt;/p></description></item><item><title>Secrets Management: Remove Credentials From Configs</title><link>https://www.netdata.cloud/blog/secrets-management/</link><pubDate>Mon, 20 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/secrets-management/</guid><description>&lt;p>If you&amp;rsquo;re running Netdata collectors that connect to databases, APIs, or other authenticated services, there&amp;rsquo;s a good chance you have passwords sitting in plain-text configuration files right now. It works, but it&amp;rsquo;s the kind of thing that makes security teams nervous and makes credential rotation painful. Every password change means editing config files and restarting collectors.&lt;/p>
&lt;p>Netdata now supports secrets management natively. Instead of putting credentials directly in your collector configurations, you reference them using a resolver syntax, and Netdata resolves the actual values at runtime from whatever source you choose: environment variables, files on disk, the output of a command, or a centralized secret store like HashiCorp Vault or AWS Secrets Manager.&lt;/p></description></item><item><title>Smarter Alerts: Test, Review &amp; Preview Schedules</title><link>https://www.netdata.cloud/blog/smarter-alerts/</link><pubDate>Thu, 16 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/smarter-alerts/</guid><description>&lt;p>Alert fatigue usually isn&amp;rsquo;t caused by one thing. It&amp;rsquo;s the accumulation of thresholds that are slightly too sensitive, alerts that fire during known maintenance windows, and historical patterns that nobody has the tools to review easily. Fixing it requires better visibility into how alerts actually behave over time, and a way to test changes before they hit production.&lt;/p>
&lt;p>We&amp;rsquo;ve shipped three improvements to alerting in Netdata that address different parts of this problem: the ability to evaluate alert definitions against historical data before deploying them, a timeline view of alert transitions, and a schedule preview for recurring silencing rules.&lt;/p></description></item><item><title>TV Mode: Put Your Dashboards on the Big Screen</title><link>https://www.netdata.cloud/blog/tv-mode/</link><pubDate>Tue, 14 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/tv-mode/</guid><description>&lt;p>One of the most common requests we&amp;rsquo;ve gotten since launching custom dashboards is deceptively simple: &amp;ldquo;How do I put this on a TV?&amp;rdquo; Teams want their dashboards on wall-mounted screens in NOCs, war rooms, and open office spaces. The dashboard is already built. The data is already there. They just need a way to display it on a screen that nobody is logged into, without exposing the full Netdata Cloud interface.&lt;/p></description></item><item><title>New Custom Dashboards: Metrics, Logs &amp; Live Commands</title><link>https://www.netdata.cloud/blog/new-custom-dashboards/</link><pubDate>Sun, 12 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/new-custom-dashboards/</guid><description>&lt;p>Custom dashboards in Netdata have always let you pull charts together on-the-fly into a single view. That&amp;rsquo;s useful, but it&amp;rsquo;s also limited. In practice, when you&amp;rsquo;re running an incident or reviewing a service, you don&amp;rsquo;t just want charts. You want to see the output of &lt;code>top&lt;/code> alongside your CPU metrics. You want slow query logs next to your database latency charts. You want an infrastructure summary card that tells you how many nodes in a room are healthy without having to click through to find out.&lt;/p></description></item><item><title>Alert Acknowledgement: Mark It as Seen, Keep Working</title><link>https://www.netdata.cloud/blog/alert-acknowledge/</link><pubDate>Fri, 10 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/alert-acknowledge/</guid><description>&lt;p>If you&amp;rsquo;ve ever opened the alerts tab during a busy period, you know the problem. There are alerts you&amp;rsquo;ve already looked at, alerts someone on your team is handling, and alerts that fired on a known issue that&amp;rsquo;s being worked on. They all sit together in the same list alongside the new ones you haven&amp;rsquo;t seen yet. There&amp;rsquo;s no way to say &amp;ldquo;I&amp;rsquo;ve seen this, move on&amp;rdquo; without silencing or disabling the alert entirely, which is a much heavier action than the situation calls for.&lt;/p></description></item><item><title>Expanded Chart View: Investigate Without Leaving the Chart</title><link>https://www.netdata.cloud/blog/charts-expanded-view/</link><pubDate>Wed, 08 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/charts-expanded-view/</guid><description>&lt;p>Charts in Netdata have always been interactive. You can zoom, pan, select time ranges, and see per-second granularity across thousands of metrics. But when you spotted something interesting, the next steps usually meant leaving the chart: opening another tab to check a related metric, navigating to the correlation tool, or pulling up a different time range for comparison. The investigation workflow lived outside the chart, even though the chart was where the investigation started.&lt;/p></description></item><item><title>Conversations: Ask Netdata About Anything You're Looking At</title><link>https://www.netdata.cloud/blog/converse-with-everything-in-netdata/</link><pubDate>Thu, 02 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/converse-with-everything-in-netdata/</guid><description>&lt;p>Netdata AI can already troubleshoot your alerts and generate Insights reports. What it couldn&amp;rsquo;t do, until now, was have a back-and-forth conversation. You could get a one-shot analysis, but you couldn&amp;rsquo;t ask follow-up questions, pull in additional context, or go from a quick question to a full investigation without starting over.&lt;/p>
&lt;p>We&amp;rsquo;ve added a conversational layer to Netdata AI. You&amp;rsquo;ll notice a new blue chat icon throughout Netdata Cloud, on charts, in the alerts table, on Insights reports, and in the reports list. Click it, and you&amp;rsquo;re in a conversation where the thing you clicked on is already the context. No copy-pasting metric names, no explaining what you&amp;rsquo;re looking at.&lt;/p></description></item><item><title>Node Groups: Organize Infrastructure Into Views</title><link>https://www.netdata.cloud/blog/node-groups/</link><pubDate>Wed, 01 Apr 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/node-groups/</guid><description>&lt;p>When you&amp;rsquo;re managing a handful of nodes, the flat list in the nodes tab works fine. When you&amp;rsquo;re managing hundreds or thousands, it becomes a wall of hostnames. You end up applying the same filters repeatedly: all the production database servers, all the nodes in eu-west, all the Kubernetes workers in the staging cluster. The filters work, but they don&amp;rsquo;t persist, and there&amp;rsquo;s no way to share them with the rest of your team.&lt;/p></description></item><item><title>Introducing the Netdata Cloud MCP Server</title><link>https://www.netdata.cloud/blog/netdata-cloud-mcp-server/</link><pubDate>Fri, 27 Feb 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-cloud-mcp-server/</guid><description>&lt;p>The Netdata Cloud MCP Server is now available — giving AI agents and assistants direct access to your Netdata through a single endpoint at &lt;code>app.netdata.cloud/api/v1/mcp&lt;/code>.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="ai-is-changing-how-we-monitor-infrastructure">AI Is Changing How We Monitor Infrastructure&lt;/h2>
&lt;p>If you&amp;rsquo;re an engineer in 2026, chances are AI is already part of your daily workflow, whether that&amp;rsquo;s a general-purpose assistant like ChatGPT, Claude, or Gemini that you bounce questions off, or a coding agent like &lt;a href="https://docs.anthropic.com/en/docs/claude-code">Claude Code&lt;/a>, &lt;a href="https://openai.com/index/codex/">Codex&lt;/a>, &lt;a href="https://www.cursor.com/">Cursor&lt;/a>, or &lt;a href="https://windsurf.com/">Windsurf&lt;/a> that writes and debugs code alongside you. These tools are incredibly powerful, but until now, they&amp;rsquo;ve been blind to what&amp;rsquo;s actually happening on your infrastructure.&lt;/p></description></item><item><title>Howard Conference &amp; Expo 2026: Smarter Observability</title><link>https://www.netdata.cloud/blog/howard-expo-2026/</link><pubDate>Tue, 03 Feb 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/howard-expo-2026/</guid><description>&lt;p>The Netdata team will be at the &lt;strong>Howard Conference and Expo &amp;ldquo;Game On&amp;rdquo;&lt;/strong> event, &lt;strong>February 24-26, 2026 at the Grand Hotel Marriott Resort in Fairhope, Alabama&lt;/strong>. We&amp;rsquo;re looking forward to meeting IT leaders and practitioners to talk about real-time observability—what it actually looks like in practice, and where traditional monitoring falls short.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-well-be-showing">What We&amp;rsquo;ll Be Showing&lt;/h2>
&lt;p>Stop by our booth to see Netdata in action and chat with our team about what you&amp;rsquo;re dealing with in your own environment.&lt;/p></description></item><item><title>India DevOps Show 2026: Modern Observability Recap</title><link>https://www.netdata.cloud/blog/india-devops-show-2026/</link><pubDate>Tue, 03 Feb 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/india-devops-show-2026/</guid><description>&lt;p>DevOps has fundamentally transformed how organizations build and deliver software. But as deployment velocity increases and infrastructure becomes more dynamic, the gap between shipping code and truly understanding system behavior continues to widen. Teams need observability that keeps pace with their pipelines, not tools that slow them down or break the budget.&lt;/p>
&lt;p>Netdata is proud to participate as a &lt;strong>Silver Partner&lt;/strong> at the &lt;strong>10th Edition India DevOps Show 2026&lt;/strong>, taking place on &lt;strong>February 13, 2026 at Aloft ORR Hotel, Bengaluru&lt;/strong>. We&amp;rsquo;re excited to engage with India&amp;rsquo;s vibrant DevOps community and share our vision for efficient, intelligent observability.&lt;/p></description></item><item><title>Tech Show London 2026: Cloud &amp; AI Observability Recap</title><link>https://www.netdata.cloud/blog/techshow-london-2026/</link><pubDate>Tue, 03 Feb 2026 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/techshow-london-2026/</guid><description>&lt;p>The intersection of cloud and AI is creating unprecedented infrastructure complexity. As organizations race to deploy AI workloads alongside traditional cloud services, the demand for intelligent, high-fidelity observability has never been greater. Understanding what&amp;rsquo;s happening across your entire stack, in real time, is no longer a luxury, it&amp;rsquo;s a necessity.&lt;/p>
&lt;p>That&amp;rsquo;s why the Netdata team is excited to be part of &lt;strong>Tech Show London 2025&lt;/strong>, taking place &lt;strong>March 4-5 at ExCeL London&lt;/strong>. We&amp;rsquo;ll be in the &lt;strong>Cloud &amp;amp; AI Infrastructure&lt;/strong> zone, ready to show you how modern observability should work.&lt;/p></description></item><item><title>Introducing Real-Time Conversations with Netdata AI</title><link>https://www.netdata.cloud/blog/ai-conversations/</link><pubDate>Tue, 23 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/ai-conversations/</guid><description>&lt;p>Over the past few months, we&amp;rsquo;ve seen incredible adoption of our AI Investigations and Insights reports. Teams are using them to automate the deep, thoughtful analysis required for complex post-mortems, capacity planning, and performance optimization. These comprehensive reports are fantastic when you need a well-researched, shareable document.&lt;/p>
&lt;p>But what about the moments &lt;em>during&lt;/em> an investigation? What about the rapid-fire &amp;ldquo;what if&amp;rdquo; questions and the quick exploration of hypotheses that happen in the heat of the moment? For that, you need speed and interactivity. You need a partner you can have a real-time dialogue with.&lt;/p></description></item><item><title>Text-To-Alert: Create Alerts From Natural Language</title><link>https://www.netdata.cloud/blog/ai-alerts/</link><pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/ai-alerts/</guid><description>&lt;p>Netdata has an incredibly powerful alerting engine. But this can sometimes be a double-edged sword: the flexibility to build incredibly specific, intelligent alerts is immense, but mastering its syntax can feel like learning a new language. We’ve heard this from so many of you. You tell us that configuring alerts is often the steepest part of the learning curve, a task that falls to the one &amp;ldquo;Netdata expert&amp;rdquo; on the team who has spent the time digging through the documentation.&lt;/p></description></item><item><title>Monitor Everything is an Anti-Pattern!</title><link>https://www.netdata.cloud/blog/monitor-everything-is-an-anti-pattern/</link><pubDate>Mon, 01 Dec 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/monitor-everything-is-an-anti-pattern/</guid><description>&lt;p>&lt;strong>Bullshit and nonsense.&lt;/strong>&lt;/p>
&lt;p>But let&amp;rsquo;s take it from the beginning.&lt;/p>
&lt;p>The industry&amp;rsquo;s story goes something like this:&lt;/p>
&lt;blockquote>
&lt;p>&lt;em>&amp;ldquo;Monitor everything is universally recognized as an anti-pattern.&amp;rdquo;&lt;/em>&lt;br/>
&lt;em>&amp;ldquo;You&amp;rsquo;ll drown in metrics, burn out your engineers, and blow your budget.&amp;rdquo;&lt;/em>&lt;br/>
&lt;em>&amp;ldquo;Just focus on 3–10 signals — the Four Golden Signals, RED, USE — and ignore everything else.&amp;rdquo;&lt;/em>&lt;br/>
&lt;em>&amp;ldquo;Trust us, you don&amp;rsquo;t want that much telemetry.&amp;rdquo;&lt;/em>&lt;br/>
&lt;br/>
(&lt;a href="https://www.netdata.cloud/resources/research/monitor-everything-anti-pattern/">true, read the whole story here&lt;/a>)&lt;/p>
&lt;/blockquote>
&lt;p>Then, in the same breath:&lt;/p></description></item><item><title>Gartner IOCS 2025: Tackling Observability Overspend</title><link>https://www.netdata.cloud/blog/gartner-iocs-2025/</link><pubDate>Fri, 07 Nov 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/gartner-iocs-2025/</guid><description>&lt;p>The observability market is facing a paradox. As organizations spend more than ever on monitoring tools, their infrastructure complexity continues to grow, and incident resolution times often remain stubbornly high. Teams are drowning in data, struggling with tool sprawl, and facing unpredictable, budget-breaking bills.&lt;/p>
&lt;p>This challenge, how to gain better visibility without spiraling costs, is one of the most critical conversations for IT leaders today.&lt;/p>
&lt;!--truncate-->
&lt;p>That&amp;rsquo;s why the Netdata team is heading to Las Vegas for the &lt;strong>Gartner IT Infrastructure, Operations &amp;amp; Cloud Strategies (IOCS) Conference&lt;/strong> from December 9-11, 2025. We&amp;rsquo;ll be there to discuss this challenge head-on and share our vision for a more efficient and intelligent future for observability.&lt;/p></description></item><item><title>ServiceNow Integration: Streamline Incident Response</title><link>https://www.netdata.cloud/blog/servicenow-integration/</link><pubDate>Fri, 07 Nov 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/servicenow-integration/</guid><description>&lt;p>When a critical alert fires at 2 AM, the last thing your on-call engineer should be doing is manual administrative work. Yet, for many teams, that&amp;rsquo;s exactly what happens. You see the alert in your monitoring tool, then you have to switch contexts, open a new browser tab, log into your ITSM platform, and manually create an incident—all while your systems are failing.&lt;/p>
&lt;p>This &amp;ldquo;swivel-chairing&amp;rdquo; between tools is slow, error-prone, and a significant drag on your Mean Time to Resolution (MTTR).&lt;/p></description></item><item><title>SOC 2 Type 2: Validated Security Controls Over Time</title><link>https://www.netdata.cloud/blog/soc2-type-2-compliance/</link><pubDate>Thu, 16 Oct 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/soc2-type-2-compliance/</guid><description>&lt;p>We&amp;rsquo;re excited to share that Netdata has successfully achieved SOC 2 Type 2 attestation.&lt;/p>
&lt;!--truncate-->
&lt;p>Following a five-month audit conducted by Sensiba LLP, we can now confirm that our security controls work consistently in practice. The audit covered the period from April 1 to August 31, 2025, and tested whether our controls operated effectively throughout that entire timeframe.&lt;/p>
&lt;p>Back in April, we announced our &lt;a href="https://www.netdata.cloud/blog/soc-2-type1/">SOC 2 Type 1 attestation&lt;/a>, which validated that our security controls were properly designed at a specific point in time. We also mentioned we were entering the monitoring period for Type 2. Today we can share the results.&lt;/p></description></item><item><title>Automate Infrastructure Analysis With AI Reports</title><link>https://www.netdata.cloud/blog/scheduled-reports-insights-investigations/</link><pubDate>Tue, 23 Sep 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/scheduled-reports-insights-investigations/</guid><description>&lt;p>The least exciting part of an operations or SRE role is often the manual, repetitive task of generating reports. It’s the Monday morning scramble to summarize weekly infrastructure health for the team, or the end-of-quarter push to build a capacity planning document. This is boilerplate work that pulls you away from critical engineering tasks.&lt;/p>
&lt;p>We believe that if a process is repeatable, it should be automated.&lt;/p>
&lt;!--truncate-->
&lt;p>That&amp;rsquo;s why we’re introducing &lt;strong>Scheduled AI Investigations and Insights&lt;/strong>. This new capability builds directly on our existing AI tools, allowing you to set your most important analyses on a recurring schedule. It’s like setting up a cron job for your infrastructure reporting, letting your Co-SRE do the heavy lifting for you.&lt;/p></description></item><item><title>AI Troubleshooting GA With On-Demand Credits</title><link>https://www.netdata.cloud/blog/ai-credits/</link><pubDate>Tue, 02 Sep 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/ai-credits/</guid><description>&lt;p>Since launching our AI investigations and insights in a research preview, one thing has become clear: &lt;strong>automated root cause analysis delivers a significant return on investment.&lt;/strong> Teams have confirmed that instant insights don&amp;rsquo;t just save a few minutes; they fundamentally shorten incident response cycles, free up valuable engineering hours, and reduce the business impact of downtime.&lt;/p>
&lt;p>The preview successfully demonstrated this value, with 10 free AI sessions per month allowing teams to integrate AI into their workflows. Now, based on the success and maturity of the capabilities, we are proud to announce that &lt;strong>Netdata&amp;rsquo;s AI investigations and insights are graduating from research preview to General Availability.&lt;/strong>&lt;/p></description></item><item><title>Save Hours on Troubleshooting with Automated Investigations</title><link>https://www.netdata.cloud/blog/automated-investigations/</link><pubDate>Mon, 04 Aug 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/automated-investigations/</guid><description>&lt;p>How many times has your team stared at a dashboard, pointed to a spike, and asked a question that charts alone can&amp;rsquo;t answer? &amp;ldquo;What was the real impact of that deployment?&amp;rdquo; &amp;ldquo;Why are our Kubernetes pods in the us-east-1 cluster suddenly crashing?&amp;rdquo; &amp;ldquo;Are we wasting money on overprovisioned servers?&amp;rdquo;&lt;/p>
&lt;!--truncate-->
&lt;p>Answering these questions is the real work of operations and SRE. It often kicks off a time-consuming scramble, sending engineers down rabbit holes for hours, days, or even weeks. You dig through logs, correlate metrics across services, and piece together clues from Slack conversations and Jira tickets.&lt;/p></description></item><item><title>Netdata Now Troubleshoots Your Alerts for You</title><link>https://www.netdata.cloud/blog/automated-alert-troubleshooting/</link><pubDate>Sun, 03 Aug 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/automated-alert-troubleshooting/</guid><description>&lt;p>The 2 AM pager alert. For anyone in Ops, SRE, or IT administration, those words trigger a familiar sense of dread. An alert has fired. Is it a real fire, or another false alarm waking you from a dead sleep? The pressure is on. Every minute of downtime costs money and reputation, but troubleshooting a complex system when you&amp;rsquo;re sleep-deprived is a Herculean task.&lt;/p>
&lt;!--truncate-->
&lt;p>This cycle is a massive drain on engineering resources. The daily grind of sifting through alerts, trying to distinguish signal from noise, and manually correlating metrics to find a root cause consumes countless hours. This constant firefighting leads to alert fatigue, where even critical notifications start to get ignored. The core questions are always the same: Is this a real problem? What is the potential impact? Why did this trigger? What do I do next? Answering them is a slow, manual, and often stressful process.&lt;/p></description></item><item><title>Netdata Implements MCP Protocol</title><link>https://www.netdata.cloud/blog/netdata-mcp-server/</link><pubDate>Wed, 18 Jun 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-mcp-server/</guid><description>&lt;p>&lt;strong>Update (February 2026):&lt;/strong> Netdata now also provides MCP via Netdata Cloud at &lt;code>app.netdata.cloud/api/v1/mcp&lt;/code> for infrastructure-wide access (Business/Homelab plan). See &lt;a href="https://learn.netdata.cloud/docs/netdata-ai/mcp">MCP documentation&lt;/a> for details.&lt;/p>
&lt;hr>
&lt;p>We are excited to announce that Netdata has officially implemented the Model Context Protocol (MCP), joining the forefront of AI-powered infrastructure monitoring.&lt;/p>
&lt;p>By enabling direct connections between AI assistants and your data sources and tools, the Model Context Protocol is a new open standard that builds a crucial link between artificial intelligence and practical systems. Instead of the AI providing generic suggestions that don’t understand your environment, MCP enables it to communicate and then interact with your infrastructure data in real time.
Being among the first monitoring platforms to adopt this groundbreaking protocol, Netdata is at the forefront of intelligent observability. With the help of this integration, traditional monitoring becomes a dialogue that your entire technical team is able to participate in.
We encourage you to dive deeper into the technical foundations of MCP, explore &lt;a href="https://www.anthropic.com/news/model-context-protocol">Anthropic&amp;rsquo;s comprehensive&lt;/a> explanation, and discover how this protocol is reshaping the future of AI-data interaction.&lt;/p></description></item><item><title>Introducing Netdata Insights</title><link>https://www.netdata.cloud/blog/netdata-insights/</link><pubDate>Tue, 27 May 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-insights/</guid><description>&lt;p>We&amp;rsquo;ve been thinking a lot about synthesis lately.&lt;/p>
&lt;p>Netdata already samples every metric every second at the edge. Engineers told us the remaining pain point was synthesis, the ability to pull hours or days or months of high‑resolution time‑series into a concise explanation they could hand to a teammate (or use themselves to debug faster).&lt;/p>
&lt;!--truncate-->
&lt;p>You know the pattern. An incident happens, and suddenly you&amp;rsquo;re context-switching between dozens of dashboards, trying to reconstruct a timeline. Or you need to write a capacity planning report, and you&amp;rsquo;re copy-pasting screenshots into slides, manually correlating trends across different retention windows. The raw data is there, but the synthesis step (the part where you turn metrics into narrative) doesn&amp;rsquo;t scale.&lt;/p></description></item><item><title>SOC 2 Type 1: Committed To Security &amp; Trust</title><link>https://www.netdata.cloud/blog/soc2-type1/</link><pubDate>Fri, 16 May 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/soc2-type1/</guid><description>&lt;p>We are pleased to announce that Netdata has successfully achieved SOC 2 Type 1 attestation!&lt;/p>
&lt;!--truncate-->
&lt;p>Following an independent examination performed by AssuranceLab CPAs LLC, the report confirms that—as of April 25, 2025—the design of Netdata’s controls meets the Security, Availability, and Confidentiality Trust Services Criteria defined by the AICPA.&lt;/p>
&lt;p>At Netdata, the security and integrity of the monitoring data our users entrust to us are paramount. This significant milestone, validated through a rigorous, independent third-party audit conducted by AssuranceLab, formally attests to the robustness of our security controls and practices as designed and implemented at a specific point in time.&lt;/p></description></item><item><title>Monitoring Netdata Restarts: A Reliable Solution</title><link>https://www.netdata.cloud/blog/2025-03-06-monitoring-netdata-restarts/</link><pubDate>Thu, 06 Mar 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/2025-03-06-monitoring-netdata-restarts/</guid><description>&lt;p>For a tool like Netdata, monitoring crashes and abnormal events extends far beyond bug fixing—it&amp;rsquo;s essential for identifying edge cases, preventing regressions, and delivering the most dependable observability experience possible. With millions of daily downloads, each event provides a vital signal for maintaining the integrity of our systems.&lt;/p>
&lt;h2 id="the-challenge-with-traditional-solutions">The Challenge with Traditional Solutions&lt;/h2>
&lt;p>Over the years, we&amp;rsquo;ve evaluated many monitoring tools, each with significant limitations:&lt;/p>
&lt;table>
 &lt;thead>
 &lt;tr>
 &lt;th>Tool&lt;/th>
 &lt;th>Strengths&lt;/th>
 &lt;th>Limitations&lt;/th>
 &lt;/tr>
 &lt;/thead>
 &lt;tbody>
 &lt;tr>
 &lt;td>&lt;strong>Sentry&lt;/strong>&lt;/td>
 &lt;td>• Comprehensive error tracking features&lt;br/>• Detailed stack traces&lt;/td>
 &lt;td>• Per-event pricing model becomes prohibitive at scale&lt;br/>• Forces sampling which reduces visibility into critical issues&lt;br/>• Compromises complete error capture for cost control&lt;/td>
 &lt;/tr>
 &lt;tr>
 &lt;td>&lt;strong>&lt;a href="https://www.netdata.cloud/solutions/technologies/gcp-monitoring/">GCP&lt;/a> BigQuery &amp;amp; Similar&lt;/strong>&lt;/td>
 &lt;td>• Powerful query capabilities&lt;br/>• Flexible data processing&lt;br/>• High scalability potential&lt;/td>
 &lt;td>• Complex reporting setup and maintenance&lt;br/>• Significant costs at high event volumes&lt;br/>• Requires specialized technical expertise&lt;/td>
 &lt;/tr>
 &lt;tr>
 &lt;td>&lt;strong>Other Solutions&lt;/strong>&lt;/td>
 &lt;td>• Various specialized features&lt;br/>• Some open-source flexibility&lt;/td>
 &lt;td>• Either too inflexible for custom requirements&lt;br/>• Or prohibitively expensive at full-capture scale&lt;br/>• Often require compromising between detail and cost&lt;/td>
 &lt;/tr>
 &lt;/tbody>
&lt;/table>
&lt;p>We consistently encountered these core challenges:&lt;/p></description></item><item><title>Netdata vs Prometheus: A 2025 Performance Analysis</title><link>https://www.netdata.cloud/blog/netdata-vs-prometheus-2025/</link><pubDate>Thu, 23 Jan 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-vs-prometheus-2025/</guid><description>&lt;p>When it comes to infrastructure monitoring, performance, scalability, and efficiency are critical considerations. In this blog post, we revisit two widely adopted open-source monitoring solutions: &lt;strong>Netdata&lt;/strong> and &lt;strong>Prometheus&lt;/strong>. Both tools have introduced notable improvements in their latest versions, emphasizing scalability and enhanced efficiency.&lt;/p>
&lt;p>In our previous &lt;a href="https://www.netdata.cloud/blog/netdata-vs-prometheus-performance-analysis/">analysis&lt;/a>, we explored key differences between these systems, focusing on resource consumption and data retention. This follow-up expands on that comparison by subjecting both tools to a significantly larger workload. With the number of monitored nodes increased to 1000, containers to 80k, and metrics ingestion reaching 4.6 million metrics per second, we examine how each system performs under these demanding conditions, focusing on CPU utilization, memory requirements, disk I/O, network usage, and data retention during data ingestion.&lt;/p></description></item><item><title>Long-Term Data Storage and Retention in Netdata</title><link>https://www.netdata.cloud/blog/long-term-data-retention/</link><pubDate>Tue, 21 Jan 2025 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/long-term-data-retention/</guid><description>&lt;p>Netdata&amp;rsquo;s database engine (dbengine) provides a sophisticated multi-tiered storage system designed for efficient long-term data retention while maintaining high granularity. This article explores the technical details of how Netdata handles metric storage, the advantages of its distributed architecture, and how to configure it for your specific needs.&lt;/p>
&lt;hr>
&lt;h2 id="database-engine-architecture">Database Engine Architecture&lt;/h2>
&lt;p>Netdata&amp;rsquo;s database engine provides a sophisticated, efficient solution for long-term metric storage through:&lt;/p>
&lt;ul>
&lt;li>Intelligent multi-tiered storage architecture&lt;/li>
&lt;li>Efficient compression and caching mechanisms&lt;/li>
&lt;li>Flexible retention strategies&lt;/li>
&lt;li>Distributed deployment options&lt;/li>
&lt;/ul>
&lt;p>This design enables organizations to maintain detailed historical data while optimizing storage use and maintaining query performance.&lt;/p></description></item><item><title>Best Of Category Badges Earned In 2024: G2 &amp; Capterra</title><link>https://www.netdata.cloud/blog/netdata-best-of-category-badges-in-2024/</link><pubDate>Mon, 23 Dec 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-best-of-category-badges-in-2024/</guid><description>&lt;h2 id="netdata-featured-with-multiple-best-of-category-badges-in-2024">Netdata Featured with Multiple “Best Of” Category Badges in 2024&lt;/h2>
&lt;p>As we are close to the end of this year, we are thrilled to announce that &lt;a href="https://www.capterra.com/p/251845/Netdata/?utm_source=vp&amp;utm_medium=blog&amp;utm_campaign=ts-q4-2024" target="_blank">&lt;strong>Netdata&lt;/strong>&lt;/a> has been recognized with multiple “Best of” badges from Gartner Digital Markets brands: &lt;a href="https://www.capterra.com/?utm_source=vp&amp;utm_medium=blog&amp;utm_campaign=ts-q4-2024" target="_blank">&lt;strong>Capterra&lt;/strong>&lt;/a>, &lt;a href="https://www.softwareadvice.com/?utm_source=vp&amp;utm_medium=blog&amp;utm_campaign=ts-q4-2024" target="_blank">&lt;strong>Software Advice&lt;/strong>&lt;/a>, and &lt;a href="https://www.getapp.com/?utm_source=vp&amp;utm_medium=blog&amp;utm_campaign=ts-q4-2024" target="_blank">&lt;strong>GetApp&lt;/strong>&lt;/a>, leading software recommendation search engines.&lt;/p>
&lt;p>This “Best of” badges program is an independent assessment that evaluates user reviews to help buyers identify the highest-rated software companies in specific categories that offer the most popular solutions.&lt;/p></description></item><item><title>Getting Started With Netdata: Real-Time Monitoring</title><link>https://www.netdata.cloud/blog/getting-started-with-netdata-a-comprehensive-guide-to-real-time-monitoring/</link><pubDate>Thu, 19 Dec 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/getting-started-with-netdata-a-comprehensive-guide-to-real-time-monitoring/</guid><description>&lt;p>Now you can start monitoring thousands of metrics in real-time, detecting anomalies throughout your infra, and troubleshooting issues even mid-crisis!
Watch this &lt;a href="https://www.youtube.com/watch?v=z5m8JdwMOn8&amp;themeRefresh=1" target="_blank">1-minute video&lt;/a> for a quick intro! Get ready to be blown away!&lt;/p>
&lt;p>To fully utilize Netdata for monitoring your infrastructure, do these:&lt;/p>
&lt;h2 id="a-hrefhttpslearnnetdataclouddocsdeployment-guides-target_blank-deploy-netdataa">&lt;a href="https://learn.netdata.cloud/docs/deployment-guides/" target="_blank">① Deploy Netdata&lt;/a>&lt;/h2>
&lt;p>&lt;em>Install Netdata to your systems.&lt;/em>&lt;/p>
&lt;p>Netdata runs on &lt;a href="https://www.netdata.cloud/solutions/technologies/linux-monitoring/">Linux&lt;/a>, &lt;a href="https://www.netdata.cloud/solutions/technologies/windows-monitoring/">Windows&lt;/a>, FreeBSD and MacOS and can be installed on &lt;a href="https://www.netdata.cloud/academy/bare-metal-server/">bare-metal servers&lt;/a>, cloud VMs, Kubernetes, even weak &lt;a href="https://www.netdata.cloud/solutions/iot-monitoring/">IoT devices&lt;/a>.
We have carefully optimized it to be the fastest and most advanced monitoring you will ever need, and at the same time be extremely friendly, polite and respectful to your production systems and applications.&lt;/p></description></item><item><title>Customize Your Netdata Experience with Favorites</title><link>https://www.netdata.cloud/blog/customize-your-netdata-experience-with-favorites/</link><pubDate>Mon, 09 Dec 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/customize-your-netdata-experience-with-favorites/</guid><description>&lt;h3 id="make-netdata-yours-with-favorites">Make Netdata yours with Favorites&lt;/h3>
&lt;p>Monitoring is personal. Different systems, workloads, and use-cases mean different charts matter to different users. Netdata now makes it easier than ever to customize your experience by letting you &lt;strong>favorite charts&lt;/strong> you care about most.&lt;/p>
&lt;h3 id="how-it-works">How it works&lt;/h3>
&lt;p>You&amp;rsquo;ll now see a &lt;strong>heart icon&lt;/strong> next to any chart or section of charts. Click it, and it instantly gets added to your &lt;strong>favorites&lt;/strong>. favorites appear right above your system overview metrics, so you see what matters most—first.&lt;/p></description></item><item><title>20 Best DevOps, SRE &amp; Observability Conferences 2025</title><link>https://www.netdata.cloud/blog/20-devops-sre-observability-events-and-conferences-you-should-consider-in-2025/</link><pubDate>Wed, 04 Dec 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/20-devops-sre-observability-events-and-conferences-you-should-consider-in-2025/</guid><description>&lt;p>Every year, DevOps, SRE Sysadmin &amp;amp; IT leaders gather at conferences across the world to share not only knowledge but also the latest trends, tools, and strategies for achieving greater insight and control over complex systems. This guide brings you a comprehensive list of DevOps, SRE &amp;amp; observability events that can help you stay ahead in a rapidly evolving field. Whether you&amp;rsquo;re exploring the latest in tools, cloud observability, or AI-driven insights, this guide will lead you to the perfect event to expand your knowledge and network with industry leaders!&lt;/p></description></item><item><title>Native Windows Agent: Real-Time Windows Monitoring</title><link>https://www.netdata.cloud/blog/netdata-native-windows-agent/</link><pubDate>Fri, 08 Nov 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-native-windows-agent/</guid><description>&lt;p>We are pleased to announce a significant advancement in system monitoring: the launch of &lt;a href="https://www.netdata.cloud/solutions/windows-monitoring/">Netdata&amp;rsquo;s first-ever Native Windows Agent&lt;/a>. This release represents a major step forward in our mission to provide comprehensive and efficient monitoring solutions across all platforms. With the introduction of the native Windows agent, we are extending our robust monitoring capabilities to Windows environments, enabling seamless and unified monitoring across diverse infrastructures. This development is a direct response to the needs of our user community, who have expressed a strong demand for a powerful and intuitive Windows monitoring solution.&lt;/p></description></item><item><title>Accurate Process Monitoring with Netdata</title><link>https://www.netdata.cloud/blog/accurate-process-monitoring-with-netdata/</link><pubDate>Mon, 04 Nov 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/accurate-process-monitoring-with-netdata/</guid><description>&lt;p>Understand why tracking cumulative resource consumption is crucial for accurate process monitoring.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="accurate-process-monitoring-with-netdata-why-tracking-cumulative-resource-consumption-matters">Accurate Process Monitoring with Netdata: Why Tracking Cumulative Resource Consumption Matters&lt;/h2>
&lt;p>Tracking the cumulative resource consumption of processes, including short-lived and exited children, is a rare feature in monitoring tools – and it’s one of the standout capabilities Netdata offers.&lt;/p>
&lt;p>Most mainstream solutions, like Datadog’s process monitoring and Prometheus’s Node Exporter, focus on active processes and only collect metrics per PID. Even specialized process monitors (&lt;code>top&lt;/code>, &lt;code>htop&lt;/code>, etc.) face the same limitations. They capture snapshots of currently running processes and often rely on fixed sampling intervals, which are too slow to catch very short-lived tasks. This approach falls short when trying to capture the resource footprint of dynamic applications, shell scripts, and complex process hierarchies.&lt;/p></description></item><item><title>Linux Load Average Myths and Realities</title><link>https://www.netdata.cloud/blog/linux-load-average-myths-and-realities/</link><pubDate>Sun, 03 Nov 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/linux-load-average-myths-and-realities/</guid><description>&lt;p>When it comes to monitoring system &lt;a href="https://www.netdata.cloud/solutions/linux-monitoring/">performance on Linux&lt;/a>, the load average is one of the most referenced metrics. Displayed prominently in tools like &lt;code>top&lt;/code>, &lt;code>uptime&lt;/code>, and &lt;code>htop&lt;/code>, it&amp;rsquo;s often used as a quick gauge of system load and capacity. But how reliable is it? For complex, multi-threaded applications, load average can paint a misleading picture of actual system performance.&lt;/p>
&lt;p>In this article, we&amp;rsquo;ll dive into the myths and realities of Linux load average, using insights from Netdata’s high-frequency, high-concurrency monitoring setup. Through this journey, we&amp;rsquo;ll uncover why load average spikes can occur even under steady workloads, and why a single metric is rarely enough to capture the true state of a system. Whether you&amp;rsquo;re a system administrator, developer, or performance enthusiast, this exploration of load average will help you interpret it more accurately and understand when it may—or may not—reflect reality.&lt;/p></description></item><item><title>12 Benefits You Get by Scaling with Netdata</title><link>https://www.netdata.cloud/blog/12-benefits-you-get-by-scaling-with-netdata/</link><pubDate>Wed, 23 Oct 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/12-benefits-you-get-by-scaling-with-netdata/</guid><description>&lt;p>&lt;a href="https://blogs.idc.com/2022/12/09/idc-futurescape-worldwide-future-of-digital-infrastructure-2023-predictions/" target="_blank">80% of decision-makers globally&lt;/a> acknowledge that digital infrastructure is essential for reaching business goals. However, IT infrastructure is becoming increasingly distributed and complex. Organizations are managing hundreds—even thousands—of nodes across cloud, on-premise, and edge environments.
This predicament makes effective monitoring across all systems more essential than ever. In turn, this drives the demand for valuable real-time insights through scalable &lt;a href="https://www.netdata.cloud/blog/understanding-monitoring-tools/" target="_blank">monitoring solutions&lt;/a> like Netdata​.&lt;/p>
&lt;p>If you’re managing a large infrastructure but haven’t fully embraced Netdata, now is the time to reconsider. Let&amp;rsquo;s take a look at the benefits you’ll get, below.&lt;/p></description></item><item><title>5 Best Datadog Alternatives For Monitoring &amp; Observability</title><link>https://www.netdata.cloud/blog/5-datadog-alternatives/</link><pubDate>Tue, 15 Oct 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/5-datadog-alternatives/</guid><description>&lt;p>As businesses rely more on &lt;a href="https://www.netdata.cloud/academy/what-is-infrastructure-monitoring-and-why-you-need-it/">infrastructure monitoring&lt;/a>, observability tools have become essential for keeping systems secure, responsive, and efficient. In this evolving monitoring landscape, Datadog is known as a leading analytics and monitoring tool, capable of collecting essential performance metrics from servers, databases, applications, and other IT infrastructures. But at what cost?&lt;/p>
&lt;h2 id="what-is-datadog-is-it-your-only-option">What Is Datadog? Is It Your Only Option?&lt;/h2>
&lt;p>Datadog is a cloud-based monitoring platform that helps organizations track the performance of their infrastructure, applications, and logs. It provides visibility across systems to ensure efficient operations and identify issues quickly.&lt;/p></description></item><item><title>ilert Integration: Streamline Monitoring &amp; Response</title><link>https://www.netdata.cloud/blog/netdata-integration-with-ilert-streamlining-monitoring-and-incident-response/</link><pubDate>Mon, 14 Oct 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-integration-with-ilert-streamlining-monitoring-and-incident-response/</guid><description>&lt;p>Netdata now integrates with ilert, a leading incident response platform. With this integration, the incident management features and alerting capabilities of ilert and the real-time systems monitoring provided by Netdata can be leveraged. By combining both systems, users can not only monitor their infrastructure with fine detail as never before, but also assure the responsiveness of critical alerts to the correct teams swiftly.&lt;/p>
&lt;h2 id="what-is-ilert">What is ilert?&lt;/h2>
&lt;p>&lt;a href="https://www.ilert.com/?utm_campaign=Netdata&amp;utm_source=integration&amp;utm_medium=organic" target="_blank">ilert&lt;/a> is an end-to-end platform for alerting, on-call management, and status pages, built for the new-age DevOps and SRE teams. It streamlines the entire incident management process by automating key aspects and providing powerful tools to improve efficiency, such as automated on-call duty, multi-level escalations, alert grouping, postmortem document creation, and much more. ilert integrates with monitoring systems, like Netdata, and sends alerts through multiple channels (SMS, phone, push, Slack, Microsoft Teams, etc.), ensuring that teams can promptly respond to issues before they impact service.&lt;/p></description></item><item><title>Important Changes to the Netdata Agent Dashboard</title><link>https://www.netdata.cloud/blog/important-changes-to-the-netdata-agent-dashboard/</link><pubDate>Mon, 12 Aug 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/important-changes-to-the-netdata-agent-dashboard/</guid><description>&lt;p>&lt;em>Important Notice: These changes ONLY impact users of the Netdata Agent Dashboard not connected to Netdata Cloud.&lt;/em>&lt;/p>
&lt;p>Dear Netdata Community,&lt;/p>
&lt;p>We are writing to inform you of upcoming changes to the Netdata Agent Dashboard, which will take effect in the coming weeks. This change impacts users from the soon to be released Netdata v2.0 onwards (and also on the Netdata v1.47 nightly releases).
Currently, the Open-Source Netdata Agents allow unauthorized and unlimited Agent dashboard access. From Netdata v2.0 onwards, all Netdata Dashboards (Agent and Cloud) will offer exactly the same functionality under the same policy. Netdata Agent Dashboard will use Netdata Cloud as an SSO provider, ensuring dashboard access is authenticated and validated by Netdata Cloud, users will have the option to proceed with an unauthorized local dashboard but this will no longer be the default.&lt;/p></description></item><item><title>Introducing Netdata's Dynamic Configuration Manager</title><link>https://www.netdata.cloud/blog/netdata-dynamic-configuration-manager/</link><pubDate>Wed, 05 Jun 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-dynamic-configuration-manager/</guid><description>&lt;p>We are thrilled to unveil the latest addition to the Netdata platform: the Dynamic Configuration Manager. This powerful new feature revolutionizes how you manage your monitoring and alerting configurations, making it easier and more efficient than ever before.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="key-features-of-the-dynamic-configuration-manager">&lt;strong>Key Features of the Dynamic Configuration Manager&lt;/strong>&lt;/h2>
&lt;ol>
&lt;li>Create and Modify Alerts from Every Chart
&lt;ul>
&lt;li>You can now create and modify alerts directly from any chart on your dashboard or from the dedicated Alerts tab. This streamlined process allows for quick adjustments and ensures your monitoring is always aligned with your current needs.&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>Configure Collectors from the Integrations Section on the Dashboard
&lt;ul>
&lt;li>Currently available for go.d collectors, this feature lets you configure collectors straight from the Integrations section. This means you can quickly identify what Netdata can monitor and set up your configurations in one go, without having to dig through multiple settings pages.&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>Submit Configurations to Multiple Nodes with One Click
&lt;ul>
&lt;li>Managing configurations across a large infrastructure can be time-consuming. With the Dynamic Configuration Manager, you can now submit configurations to multiple nodes simultaneously, saving you valuable time and effort.&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>Construct and Copy Configurations for IaC Solutions
&lt;ul>
&lt;li>For those using Infrastructure as Code (IaC) solutions, this feature allows you to construct and copy configurations easily, integrating them into your IaC workflows. This ensures your configurations are consistent and reproducible across different environments.&lt;/li>
&lt;/ul>
&lt;/li>
&lt;/ol>
&lt;h2 id="how-to-use-the-dynamic-configuration-manager">&lt;strong>How to Use the Dynamic Configuration Manager&lt;/strong>&lt;/h2>
&lt;p>To help you get started with the Dynamic Configuration Manager, we’ve put together a quick guide using the Netdata demo environment.&lt;/p></description></item><item><title>How to automate adding nodes to rooms in Netdata?</title><link>https://www.netdata.cloud/blog/how-can-netdata-agents-be-placed-in-different-rooms-in-an-automated-way/</link><pubDate>Tue, 28 May 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/how-can-netdata-agents-be-placed-in-different-rooms-in-an-automated-way/</guid><description>&lt;p>How we organize nodes (and the Netdata agents that are running on those nodes) across different rooms should reflect our architectural decision because the room is a logical container with its own user members and notification rules. So if we are monitoring large infrastructure we should be consistent with these rules and one way to achieve this is to choose automation. &lt;a href="https://registry.terraform.io/providers/netdata/netdata/latest">Netdata Cloud Terraform Provider&lt;/a> lets you automate this by provisioning all the cloud resources and giving you the credentials to spin up the Netdata Agents. In this article, we will concentrate on how in practice we can organize and assign nodes across different rooms in two scenarios, in each of them I&amp;rsquo;m using &lt;strong>non-production&lt;/strong> installation of the Netdata Agents:&lt;/p></description></item><item><title>Costa Tsaousis Interview: Homelab Show Insights</title><link>https://www.netdata.cloud/blog/interview-recap-with-costa-tsaousis-ceo-of-netdata-insights-from-the-homelab-show/</link><pubDate>Fri, 26 Apr 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/interview-recap-with-costa-tsaousis-ceo-of-netdata-insights-from-the-homelab-show/</guid><description>&lt;p>Costa Tsaousis, founder and chief visionary of Netdata, recently shared his expertise on real-time monitoring&amp;rsquo;s pivotal role in modern IT landscapes during an episode of ‘The Homelab Show’ on YouTube. If you missed the live stream, here’s an essential summary of the discussion’s key points.&lt;/p>
&lt;iframe width="560" height="315" src="https://www.youtube.com/embed/WvXab8MkRS4?si=QkJE_brYT3aZ1FZQ" title="YouTube video player" frameborder="0" allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" referrerpolicy="strict-origin-when-cross-origin" allowfullscreen>&lt;/iframe>
&lt;h2 id="the-necessity-of-real-time-monitoring">The Necessity of Real-Time Monitoring&lt;/h2>
&lt;p>Nowadays, data flows incessantly and operational demands are continuous, the importance of real-time monitoring cannot be overstated.&lt;/p></description></item><item><title>Decentralized Monitoring Explained</title><link>https://www.netdata.cloud/blog/decentralized-or-distributed-monitoring-explained/</link><pubDate>Thu, 11 Apr 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/decentralized-or-distributed-monitoring-explained/</guid><description>&lt;h2 id="introduction-to-distributed-observability">Introduction to Distributed Observability&lt;/h2>
&lt;p>Users often find themselves puzzled by the concepts of decentralized or distributed monitoring. This confusion is likely due to many monitoring systems claiming distributed capabilities, making it challenging to discern how Netdata stands out.&lt;/p>
&lt;p>To grasp the distinction, we must delve into the evolution of monitoring systems.&lt;/p>
&lt;p>When the first monitoring systems were created, about 20-25 years ago, they were built as SNMP collectors. The monitoring application was installed on a server, configured to discover network devices via SNMP, pulling data once every minute, per device. Simultaneously, the monitoring system was exposing a daemon to collect SNMP traps (key events generated by the network devices, pushed to the monitoring system).&lt;/p></description></item><item><title>Netdata is the only real-time monitoring solution: Justified</title><link>https://www.netdata.cloud/blog/netdata-is-the-only-real-time-monitoring-solution-justified/</link><pubDate>Wed, 10 Apr 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-is-the-only-real-time-monitoring-solution-justified/</guid><description>&lt;p>In the digital era, where data flows like a ceaseless river, real-time monitoring stands as a pivotal technology, allowing organizations to not only keep pace but also to deeply understand the intricate dance of their operational ecosystems. This technology is not just about keeping tabs; it&amp;rsquo;s about gaining a profound, almost intuitive sense of the micro-worlds within which systems, containers, services, and applications pulse and thrive.&lt;/p>
&lt;p>Real-time monitoring is the art and science of tracking system performance, activities, or transactions continuously and automatically, providing the ability to analyze and visualize data the moment it&amp;rsquo;s generated.&lt;/p></description></item><item><title>Understanding Monitoring Tools</title><link>https://www.netdata.cloud/blog/understanding-monitoring-tools/</link><pubDate>Wed, 10 Apr 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/understanding-monitoring-tools/</guid><description>&lt;p>If you care about operational excellence when it comes to your IT infrastructure, the role of monitoring systems is pivotal. As we navigate through the myriad of available monitoring tools, it becomes essential to understand the distinct architectures, styles, and focal points of various monitoring solutions, as well as the time-to-value they offer. This blog post aims to demystify the landscape of monitoring systems, providing a comprehensive overview that categorizes these tools into four primary architectural design principles.&lt;/p></description></item><item><title>SafetyDetectives: An Interview-With-Costa-Tsaousis</title><link>https://www.netdata.cloud/blog/safetydetectives/</link><pubDate>Mon, 01 Apr 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/safetydetectives/</guid><description>&lt;p>In a recent conversation with SafetyDetectives, Costa Tsaousis, CEO and founder of Netdata, shares insights into the inception and evolution of Netdata, a game-changing monitoring solution. With a background in fintech and a passion for real-time data processing, Tsaousis was driven to create Netdata in response to the significant gaps he identified in traditional monitoring tools. Emphasizing the importance of real-time data, comprehensive metrics collection, and the innovative use of machine learning, Tsaousis discusses how Netdata is setting new standards in the monitoring industry. His vision for Netdata not only challenges the status quo but also introduces a novel approach to cybersecurity, making it an essential tool for organizations worldwide.&lt;/p></description></item><item><title>Manage Netdata Cloud with Terraform</title><link>https://www.netdata.cloud/blog/netdata-terraform-provider/</link><pubDate>Wed, 27 Mar 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-terraform-provider/</guid><description>&lt;p>We proudly announce the release of the &lt;a href="https://registry.terraform.io/providers/netdata/netdata/latest">Netdata Cloud Terraform Provider&lt;/a>.&lt;/p>
&lt;p>It&amp;rsquo;s a step forward to make our platform more automated and compliant with the modern Infrastructure as Code approach. &lt;a href="https://www.terraform.io/">Terraform&lt;/a> is one of the leaders in the IaC tools with a rich ecosystem of providers and modules, now you can put a puzzle with Netdata Cloud to your stack.&lt;/p>
&lt;p>The initial iteration of the Netdata Cloud Terraform Provider supports the following resources:&lt;/p></description></item><item><title>Dynatrace vs Datadog vs Instana vs Grafana vs Netdata!</title><link>https://www.netdata.cloud/blog/netdata-vs-datadog-dynatrace-instana-grafana/</link><pubDate>Sun, 17 Mar 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-vs-datadog-dynatrace-instana-grafana/</guid><description>&lt;p>In this post, we delve into the comparative analysis of the commercial offerings of five leading monitoring solutions—Dynatrace, &lt;a href="https://www.netdata.cloud/blog/5-datadog-alternatives/">Datadog&lt;/a>, Instana, Grafana, and Netdata. Our objective is to unravel the intrinsic value each of these services offers when applied to a real-world scenario. To accomplish this, we employed trial subscriptions of these services to monitor a set of Ubuntu servers and VMs, each hosting a pair of widely-used applications: NGINX and PostgreSQL, along with a couple of Docker and LXC containers. Additionally, we extended our monitoring to physical servers to evaluate the efficacy of these tools in capturing hardware and sensor data along with VMs monitored from the host.&lt;/p></description></item><item><title>New Streamlined Plan Structure</title><link>https://www.netdata.cloud/blog/netdata-unified-plans/</link><pubDate>Wed, 06 Mar 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-unified-plans/</guid><description>&lt;blockquote>
&lt;p>&lt;strong>UPDATE:&lt;/strong> Netdata is introducing a streamlined plan structure, sunsetting Early Bird plans on 13-03-2024.&lt;/p>
&lt;/blockquote>
&lt;!--truncate-->
&lt;p>As the landscape of real-time monitoring evolves, so does the diversity and complexity of use cases that our community brings to Netdata. Our mission has always been to democratize monitoring by making it accessible, powerful, and scalable for everyone. With the rapid growth of our user base and their expanding needs, it&amp;rsquo;s become clear that our plan structure must evolve to maintain this mission sustainably.&lt;/p></description></item><item><title>Upcoming Homelab Plan</title><link>https://www.netdata.cloud/blog/netdata-homelab-plan/</link><pubDate>Wed, 07 Feb 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-homelab-plan/</guid><description>&lt;blockquote>
&lt;p>&lt;strong>UPDATE:&lt;/strong> On the 2024-02-08 A new Netdata Cloud Homelab plan will become available. This is aimed for home users and students.&lt;/p>
&lt;/blockquote>
&lt;h2 id="what-you-need-to-know">What you need to know?&lt;/h2>
&lt;ul>
&lt;li>
&lt;p>New Plan alert: We&amp;rsquo;re introducing a dedicated plan on Netdata Cloud—the Homelab plan. It&amp;rsquo;s tailored to meet the needs of home users and students, offering unrestricted access to Netdata features without the limitations seen in the Community plan.&lt;/p>
&lt;/li>
&lt;li>
&lt;p>Exclusively for Personal Use: The Homelab plan is designed for personal, non-commercial use only. To qualify, users will need to self-certify as a home user or student during the sign-up process.&lt;/p></description></item><item><title>IoT Monitoring Challenges: Key Issues &amp; How To Overcome Them</title><link>https://www.netdata.cloud/blog/iot-monitoring-challenges/</link><pubDate>Thu, 11 Jan 2024 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/iot-monitoring-challenges/</guid><description>&lt;p>With the increasing prevalence of IoT devices, which are being used in a wide range of applications, from smart homes and cities to industrial and agricultural systems, monitoring thei performance and health is extremely important. However, it’s essential to remember that monitoring IoT devices involves more than just tracking device-level data. In addition, monitoring data from the IoT platform or application layer is equally important.&lt;/p>
&lt;p>We’ll explore some of these topics in more detail and explain how Netdata can play an essential role in the monitoring of such devices, including some hints on how it can be set up for maximum performance in such scenarios.&lt;/p></description></item><item><title>Introducing Netdata’s Alerts Configuration Manager</title><link>https://www.netdata.cloud/blog/netdata-alerts-configuration-manager/</link><pubDate>Thu, 21 Dec 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-alerts-configuration-manager/</guid><description>&lt;p>Netdata introduces its latest feature, the Alerts Configuration Manager, transforming the way users configure and manage alerts in their Netdata environment. This powerful tool integrates directly into the Netdata Dashboard, offering a streamlined and intuitive interface for both novice and experienced users.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-is-the-alerts-configuration-manager">What is the Alerts Configuration Manager?&lt;/h2>
&lt;p>The Alerts Configuration Manager is an innovative feature available to users with Business subscriptions. It allows for the creation and customization of alerts directly from the Netdata Dashboard, employing a user-friendly UI wizard. This tool simplifies alert configuration, making it accessible and straightforward, even for those who are not deeply technical.&lt;/p></description></item><item><title>Cost Transparency: The True Cost Of Monitoring</title><link>https://www.netdata.cloud/blog/netdata-cost-transparency-unveiling-the-true-cost-of-monitoring/</link><pubDate>Fri, 01 Dec 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-cost-transparency-unveiling-the-true-cost-of-monitoring/</guid><description>&lt;p>Businesses are increasingly reliant on monitoring tools to ensure the seamless performance and reliability of their systems. However, the true cost of implementing and maintaining these tools is often obscured by hidden expenses. Our previous blog delved into the concealed costs associated with various monitoring solutions, such as Prometheus &amp;amp; Grafana (Open Source Monitoring) and commercial platforms like Datadog, Dynatrace, and NewRelic. These costs can manifest in various forms - from complex setups and maintenance to additional charges for advanced features.&lt;/p></description></item><item><title>Understanding the Netdata Methodology</title><link>https://www.netdata.cloud/blog/netdata-methodology/</link><pubDate>Sat, 04 Nov 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-methodology/</guid><description>&lt;p>In the dynamic landscape of modern infrastructure and multi cloud environments, observing and understanding system performance requires a new breed of tools—ones that keep pace with the &amp;rsquo;living&amp;rsquo; nature of modern infrastructure. This is the inflection point at which Netdata steps in, and aims to bring a fresh perspective to monitoring.&lt;/p>
&lt;h2 id="what-does-netdata-do-differently">What does Netdata do differently&lt;/h2>
&lt;p>There are key “cultural” faults that we believe hold the &lt;a href="https://www.netdata.cloud/blog/monitoring-vs-observability">observability&lt;/a> industry back. Let’s take a look through the prism of these faults and understand what Netdata is doing differently to address them.&lt;/p></description></item><item><title>Netdata Best Practices</title><link>https://www.netdata.cloud/blog/netdata-best-practices/</link><pubDate>Fri, 03 Nov 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-best-practices/</guid><description>&lt;p>Effective &lt;strong>system monitoring&lt;/strong> is non-negotiable in today&amp;rsquo;s complex IT environments. Netdata offers real-time performance and health monitoring with precision and granularity. But the key to harnessing its full potential lies in the optimization of your setup. Let’s ensure you are not just collecting data, but doing it in the most optimal way while gaining actionable insights from it.&lt;/p>
&lt;p>The starting point for optimization is a robust setup. Netdata is engineered for minimal footprint and can run on a wide range of hardware—from IoT devices to powerful servers. Time for a deep dive into each of these key areas and what the best practices you should follow, if you are serious about monitoring and optimizing your Netdata monitoring setup:&lt;/p></description></item><item><title>System Operators: The Role Of SysOps In IT Infrastructure</title><link>https://www.netdata.cloud/blog/system-operators-unlock-log-management-mastery-with-systemd-journal-and-netdata/</link><pubDate>Fri, 03 Nov 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/system-operators-unlock-log-management-mastery-with-systemd-journal-and-netdata/</guid><description>&lt;p>System operators know the drill: as the complexity of systems scales, so does the deluge of logs. Traditionally, taming this relentless tide demands a concoction of costly tools and laborious configurations—until now. The dynamic duo of &lt;code>systemd-journal&lt;/code> and Netdata is revolutionizing log management, turning what was once a Herculean task into a streamlined, powerful, and surprisingly straightforward process.&lt;/p>
&lt;h2 id="efficient-handling-of-volume-and-velocity">Efficient Handling of Volume and Velocity&lt;/h2>
&lt;p>&lt;code>systemd-journal&lt;/code> is built to manage the deluge of data that systems generate, &lt;strong>without buckling under the speed and volume of incoming logs&lt;/strong>. It captures logs at the source, facilitating direct and immediate processing. By utilizing the journal&amp;rsquo;s native mechanism to send logs to a central server, system operators can bypass the complexities of traditional &lt;strong>log centralization methods&lt;/strong>.&lt;/p></description></item><item><title>Upcoming Changes To Netdata Cloud Plans</title><link>https://www.netdata.cloud/blog/netdata-plan-changes/</link><pubDate>Thu, 02 Nov 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-plan-changes/</guid><description>&lt;blockquote>
&lt;p>&lt;strong>UPDATE&lt;/strong>: On the &lt;strong>2023-11-08&lt;/strong> Node and Dashboard limits will be applied on the Netdata Cloud &lt;strong>Community plan&lt;/strong>, while all current features of the Community plan will remain the same.&lt;/p>
&lt;/blockquote>
&lt;h2 id="what-you-need-to-know">What you need to know?&lt;/h2>
&lt;ul>
&lt;li>
&lt;p>On Netdata Cloud Free Community plan, the number of &lt;strong>active nodes&lt;/strong> that can be concurrently visualized on the Netdata dashboards, as well as the number of &lt;strong>active custom dashboards&lt;/strong> for accounts created after 2023-11-07 will be subject to limits.&lt;/p></description></item><item><title>Netdata vs Prometheus</title><link>https://www.netdata.cloud/blog/netdata-vs-prometheus-performance-analysis/</link><pubDate>Sat, 28 Oct 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-vs-prometheus-performance-analysis/</guid><description>&lt;p>In an era dominated by data-driven decision making, monitoring tools play an indispensable role in ensuring that our systems run efficiently and without interruption. When considering tools like &lt;strong>Netdata and Prometheus&lt;/strong>, performance isn&amp;rsquo;t just a number; it&amp;rsquo;s about empowering users with &lt;strong>real-time insights&lt;/strong> and enabling them to act with agility.&lt;/p>
&lt;p>There&amp;rsquo;s a genuine need in the community for tools that are not only comprehensive in their offerings but also &lt;strong>swift and scalable&lt;/strong>. This desire stems from our evolving digital landscape, where the ability to swiftly detect, diagnose, and &lt;a href="https://www.netdata.cloud/kubernetes-monitoring/">rectify anomalies&lt;/a> has direct implications on user experiences and business outcomes. Especially as infrastructure grows in complexity and scale, there&amp;rsquo;s an increasing demand for &lt;strong>monitoring tools&lt;/strong> to keep up and provide clear, timely insights.&lt;/p></description></item><item><title>Discover The New Netdata!</title><link>https://www.netdata.cloud/blog/discover-the-new-netdata/</link><pubDate>Fri, 27 Oct 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/discover-the-new-netdata/</guid><description>&lt;p>Missed the last &lt;strong>Netdata&lt;/strong> updates? Here is what is new:&lt;/p>
&lt;h2 id="explore-your-systemd-journal-logs-with-netdata">Explore your systemd-journal logs with Netdata&lt;/h2>
&lt;p>&lt;img src="https://github.com/netdata/blog/assets/139226121/7d2779c9-0efb-4491-8fe3-aedce1dc72fb" alt="systemd-journal-logs">&lt;/p>
&lt;p>Netdata &lt;a href="https://learn.netdata.cloud/docs/logs/systemd-journal/?utm_source=IL&amp;amp;utm_medium=internallinking&amp;amp;utm_campaign=new_netada">got a &lt;code>systemd&lt;/code>-journal logs explorer&lt;/a> to analyze your &lt;code>systemd&lt;/code>-journal logs, directly on their sources. By just installing &lt;strong>Netdata&lt;/strong> on any systemd based system, Netdata automatically finds all the &lt;strong>journal sources&lt;/strong> and presents a powerful dashboard to explore, search, filter and analyze your &lt;strong>logs&lt;/strong>. It works on both individual servers and journal centralization servers.&lt;/p></description></item><item><title>Improve Your Security With systemd-journal &amp; Netdata</title><link>https://www.netdata.cloud/blog/improve-your-security-with-systemd-and-netdata/</link><pubDate>Tue, 24 Oct 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/improve-your-security-with-systemd-and-netdata/</guid><description>&lt;p>&lt;strong>&lt;code>systemd&lt;/code> journals&lt;/strong> play a crucial role in the Linux system ecosystem, and understanding the importance of the logs contained within is essential for both system administrators and developers.&lt;/p>
&lt;p>For those unfamiliar, &lt;code>systemd&lt;/code> is an init system employed by Linux distributions, initiates the user space and oversees all ensuing processes. One of its key components, &lt;code>systemd&lt;/code> journal, assumes a central role in logging system activities and messages, delivering a host of benefits to both system administrators, developers and cyber security engineers. The &lt;code>systemd&lt;/code> journal functions as a logging system that gathers, archives, and oversees log messages and event data originating from a diverse array of system components, encompassing the kernel, system services, applications, and user activities.&lt;/p></description></item><item><title>Monitoring vs Observability: Key Differences</title><link>https://www.netdata.cloud/blog/monitoring-vs-observability/</link><pubDate>Tue, 24 Oct 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/monitoring-vs-observability/</guid><description>&lt;p>As systems increasingly shift towards distributed architectures to deliver application services, the roles of monitoring and observability have never been more crucial. Monitoring delivers the situational awareness you need to detect issues, while &lt;a href="https://www.netdata.cloud/academy/what-is-observability/">observability goes a step further&lt;/a>, offering the analytical depth to understand the root cause of those issues.&lt;/p>
&lt;p>Understanding the nuanced differences between monitoring and observability is crucial for anyone responsible for system health and performance. In dissecting these methodologies, we&amp;rsquo;ll explore their unique strengths, dive into practical applications, and illuminate how to strategically employ each to enhance operational outcomes.&lt;/p></description></item><item><title>Exploring systemd journal logs with Netdata</title><link>https://www.netdata.cloud/blog/exploring-systemd-journal-logs/</link><pubDate>Thu, 12 Oct 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/exploring-systemd-journal-logs/</guid><description>&lt;p>Today, we released our &lt;code>systemd&lt;/code> &lt;strong>journal plugin for Netdata&lt;/strong>, allowing you to explore, view, search, filter and analyze &lt;code>systemd&lt;/code> journal logs.&lt;/p>
&lt;p>Like most things about Netdata, this is a &lt;strong>zero-configuration plugin&lt;/strong>. You don’t have to do anything apart from &lt;strong>installing Netdata&lt;/strong> on your systems.This is key design direction for Netdata, since we want Netdata to be able to help even if you install it mid-crisis, while you have an incident at hand.&lt;/p></description></item><item><title>systemd journal logs</title><link>https://www.netdata.cloud/blog/systemd-journal-logs/</link><pubDate>Mon, 09 Oct 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/systemd-journal-logs/</guid><description>&lt;p>&lt;em>“Why bother with it? I let it run in the background and focus on more important DevOps work.”&lt;/em>
— a random DevOps Engineer at Reddit r/devops&lt;/p>
&lt;p>In an era where technology is evolving at breakneck speeds, it&amp;rsquo;s easy to overlook the tools that are right under our noses. One such underutilized powerhouse is the &lt;strong>&lt;code>systemd&lt;/code> journal&lt;/strong>. For many, it&amp;rsquo;s a mere tool to check the status of systemd service units or to tail the most recent events (journalctl -f). Others who do mainly container work, ignore even its existence.&lt;/p></description></item><item><title>Netdata Cloud On Prem</title><link>https://www.netdata.cloud/blog/netdata-cloud-on-prem/</link><pubDate>Tue, 26 Sep 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-cloud-on-prem/</guid><description>&lt;p>We at &lt;strong>Netdata&lt;/strong> understand that &lt;a href="https://blog.netdata.cloud/future-of-infrastructure-monitoring/">infrastructure monitoring&lt;/a> can be a complex maze—high costs, specialized skill sets, scalability, data silos, and more. That&amp;rsquo;s why we have always aimed to streamline and modernize this critical operation. Today, we&amp;rsquo;re thrilled to announce the launch of &lt;a href="https://www.netdata.cloud/contact-us/?subject=on-prem">Netdata Cloud On Prem&lt;/a>, a ground-breaking solution designed for robust &lt;strong>on-prem infrastructure monitoring&lt;/strong> - it comes with all the &lt;strong>Netdata Cloud&lt;/strong> features you love but fully on prem.&lt;/p>
&lt;h2 id="netdata-cloud-on-prem">&lt;strong>Netdata Cloud On-Prem&lt;/strong>&lt;/h2>
&lt;p>While &lt;a href="https://www.netdata.cloud/">Netdata Cloud&lt;/a> never stores any of your metric data on the cloud and just streams it ephemerally while you view a chart, the demand for on premise infrastructure monitoring has never been more pressing. Many large enterprises, governmental organizations, research institutes and critical infrastructures require a level of data privacy, security, and customization that only an on-prem solution can offer.&lt;/p></description></item><item><title>Netdata QoS Classes monitoring</title><link>https://www.netdata.cloud/blog/netdata-qos-monitoring/</link><pubDate>Tue, 26 Sep 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-qos-monitoring/</guid><description>&lt;p>Netdata monitors &lt;code>tc&lt;/code> QoS classes for all interfaces.&lt;/p>
&lt;p>If you also use &lt;a href="http://firehol.org/tutorial/fireqos-new-user/">FireQOS&lt;/a> it will collect interface and class names.&lt;/p>
&lt;p>There is a &lt;a href="https://raw.githubusercontent.com/netdata/netdata/master/collectors/tc.plugin/tc-qos-helper.sh.in">shell helper&lt;/a> for this (all parsing is done by the plugin in &lt;code>C&lt;/code> code - this shell script is just a configuration for the command to run to get &lt;code>tc&lt;/code> output).&lt;/p>
&lt;p>The source of the tc plugin is &lt;a href="https://raw.githubusercontent.com/netdata/netdata/master/collectors/tc.plugin/plugin_tc.c">here&lt;/a>. It is somewhat complex, because a state machine was needed to keep track of all the &lt;code>tc&lt;/code> classes, including the pseudo classes tc dynamically creates.&lt;/p></description></item><item><title>Netdata, Prometheus, Grafana Stack</title><link>https://www.netdata.cloud/blog/netdata-prometheus-grafana-stack/</link><pubDate>Tue, 26 Sep 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-prometheus-grafana-stack/</guid><description>&lt;p>In this blog, we will walk you through the basics of getting Netdata, Prometheus and Grafana all working together and
&lt;a href="https://www.netdata.cloud/blog/web-servers-and-their-performance/">monitoring your application servers&lt;/a>. This article will be using docker on your local workstation. We will be working
with docker in an ad-hoc way, launching containers that run &lt;code>/bin/bash&lt;/code> and attaching a TTY to them. We use docker here
in a purely academic fashion and do not condone running Netdata in a container. We pick this method so individuals
without &lt;a href="https://www.netdata.cloud/academy/what-is-cloud-management-how-to-maximize-efficiency/">cloud accounts&lt;/a> or access to VMs can try this out and for it&amp;rsquo;s speed of deployment.&lt;/p></description></item><item><title>Process Monitoring vs Console Tools: A Comparison</title><link>https://www.netdata.cloud/blog/netdata-processes-monitoring-comparison-with-console-tools/</link><pubDate>Tue, 26 Sep 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-processes-monitoring-comparison-with-console-tools/</guid><description>&lt;p>Netdata reads &lt;code>/proc/&amp;lt;pid&amp;gt;/stat&lt;/code> for all processes, once per second and extracts &lt;code>utime&lt;/code> and
&lt;code>stime&lt;/code> (user and system cpu utilization), much like all the console tools do.&lt;/p>
&lt;p>But it also extracts &lt;code>cutime&lt;/code> and &lt;code>cstime&lt;/code> that account the user and system time of the exit children of each process.
By keeping a map in memory of the whole process tree, it is capable of assigning the right time to every process, taking
into account all its exited children.&lt;/p></description></item><item><title>Our first ML based anomaly alert</title><link>https://www.netdata.cloud/blog/our-first-ml-based-anomaly-alert/</link><pubDate>Wed, 13 Sep 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/our-first-ml-based-anomaly-alert/</guid><description>&lt;p>Over the last few years we have slowly and methodically been building out the &lt;a href="https://learn.netdata.cloud/docs/ml-and-troubleshooting/">ML based capabilities&lt;/a> of the Netdata agent, dogfooding and iterating as we go. To date, these features have mostly been somewhat reactive and tools to aid once you are already troubleshooting.&lt;/p>
&lt;p>Now we feel we are ready to take a first gentle step into some more proactive use cases, starting with a &lt;a href="https://github.com/netdata/netdata/pull/14687">simple node level anomaly rate alert&lt;/a>.&lt;/p></description></item><item><title>Anomaly Rate By Type</title><link>https://www.netdata.cloud/blog/anomaly-rate-by-type/</link><pubDate>Wed, 30 Aug 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/anomaly-rate-by-type/</guid><description>&lt;p>We have &lt;a href="https://github.com/netdata/netdata/pull/15856">recently added&lt;/a> a more detailed anomaly rate chart to Netdata that breaks out the overall &lt;a href="https://learn.netdata.cloud/docs/ml-and-troubleshooting/machine-learning-ml-powered-anomaly-detection#node-anomaly-rate">node anomaly rate&lt;/a> by type, this lets you more easily see what parts of your infrastructure might be experiencing an uptick in anomalies when you see the overall node anomaly rate increase.&lt;/p>
&lt;h2 id="what-is-type">What is &lt;code>type&lt;/code>?&lt;/h2>
&lt;p>&lt;code>type&lt;/code> is generally the prefix of the chart id in Netdata and controls where charts live within the menu on the overview page, for example the &lt;code>mem.available&lt;/code> chart has a type of &lt;code>mem&lt;/code> which in part controls why it lives under the &amp;ldquo;Memory&amp;rdquo; section of the menu.&lt;/p></description></item><item><title>Release 1.41: Brand-New UI For Agents &amp; Parents</title><link>https://www.netdata.cloud/blog/netdata-version-1.41/</link><pubDate>Thu, 20 Jul 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-version-1.41/</guid><description>&lt;p>Netdata Agents and Parents now have a new UI!&lt;/p>
&lt;p>Checkout the release meetup video or read on to learn more about the new UI and other features in this release.&lt;/p>

 &lt;div style="position: relative; padding-bottom: 56.25%; height: 0; overflow: hidden;">
 &lt;iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" allowfullscreen="allowfullscreen" loading="eager" referrerpolicy="strict-origin-when-cross-origin" src="https://www.youtube.com/embed/WCUn4-LneCw?autoplay=0&amp;amp;controls=1&amp;amp;end=0&amp;amp;loop=0&amp;amp;mute=0&amp;amp;start=0" style="position: absolute; top: 0; left: 0; width: 100%; height: 100%; border:0;" title="YouTube video">&lt;/iframe>
 &lt;/div>

&lt;ul>
&lt;li>&lt;a href="#v1410-netdata-open-source-growth">Netdata Growth &lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-release-highlights">Release Highlights&lt;/a>
&lt;ul>
&lt;li>&lt;strong>&lt;a href="#v1410-one-dashboard">New Agent Dashboard!&lt;/a>&lt;/strong>&lt;/li>
&lt;li>&lt;a href="#v1410-netdata-assistant">Netdata Assistant&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-netdata-freeipmi">New FreeIPMI collector for monitoring enterprise hardware&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-netdata-apps">Netdata Detects FDs Leaking&lt;/a>&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#v1410-acknowledgements">Acknowledgements &lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-contributions">Contributions&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-contributions-collectors">Collectors&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-contributions-documentation">Documentation &lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-contributions-packaging">Packaging/Installation&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-contributions-health">Health&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-contributions-exporting">Exporting&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-contributions-other">Other Notable Changes&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-deprecation-notice">Deprecation notice&lt;/a>
&lt;ul>
&lt;li>&lt;a href="#v1410-deprected-in-this-release">Deprecated in this release&lt;/a>&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#v1410-netdata-release-meetup">Netdata Release Meetup&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1410-support-options">Support options&lt;/a>&lt;/li>
&lt;/ul>
&lt;p>Steady to our schedule, this is another great Netdata release!&lt;/p></description></item><item><title>Netdata Assistant: Your AI-Powered Troubleshooting Sidekick</title><link>https://www.netdata.cloud/blog/netdata-assistant/</link><pubDate>Fri, 14 Jul 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-assistant/</guid><description>&lt;p>Hey there! We&amp;rsquo;re excited to share a new troubleshooting feature we have added to Netdata, the Netdata Assistant. We&amp;rsquo;ve built this tool to help you troubleshoot more effectively and with less stress. Let&amp;rsquo;s dive in.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="whats-the-netdata-assistant">What&amp;rsquo;s the Netdata Assistant?&lt;/h2>
&lt;p>The Netdata Assistant is an AI tool that uses large language models and our community&amp;rsquo;s knowledge to guide you during troubleshooting.&lt;/p>
&lt;p>Here&amp;rsquo;s a scenario. It&amp;rsquo;s 3 am and you get an alert. Instead of scrambling to Google what&amp;rsquo;s going on, you can just click on the assistant button. The Netdata Assistant will give you the lowdown on the alert, why it&amp;rsquo;s happening, and why you should care. It&amp;rsquo;ll also guide you on how to troubleshoot it and even offer some handy web links for more info, if you&amp;rsquo;re interested.&lt;/p></description></item><item><title>Hidden Costs Of Monitoring: Uncovering Expenses &amp; Solutions</title><link>https://www.netdata.cloud/blog/hidden-costs-of-monitoring/</link><pubDate>Fri, 07 Jul 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/hidden-costs-of-monitoring/</guid><description>&lt;p>When it comes to monitoring IT infrastructure, the &lt;a href="https://www.netdata.cloud/pricing/">costs you see on the price tag&lt;/a> of the tool are often just the tip of the iceberg. Below the waterline, a mass of hidden costs can lurk, which can significantly affect the total cost of ownership.&lt;/p>
&lt;!--truncate-->
&lt;p>In this blogpost we will cover the analysis of two traditional monitoring domains, &lt;a href="https://www.netdata.cloud/open-source/">Open Source observability&lt;/a> and Commercial Centralized observability solutions, focusing the direct and indirect impacts when implementing these solution. In summary:&lt;/p></description></item><item><title>Netdata &amp; Ansible example: ML demo room</title><link>https://www.netdata.cloud/blog/ml-demo-ansible-configuration-management/</link><pubDate>Fri, 07 Jul 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/ml-demo-ansible-configuration-management/</guid><description>&lt;p>We are always trying to lower the barrier to entry when it comes to monitoring and observability and one place we have consistently witnessed some pain from users is around adopting and approaching &lt;a href="https://www.atlassian.com/microservices/microservices-architecture/configuration-management">configuration management&lt;/a> tools and practices as your infrastructure grows and becomes more complex.&lt;/p>
&lt;p>To that end, we have begun recently publishing our own &lt;a href="https://github.com/netdata/community/tree/main/configuration-management/ansible-ml-demo">little example ansible project&lt;/a> used to maintain and manage the servers used in our public &lt;a href="https://app.netdata.cloud/spaces/netdata-demo/rooms/machine-learning/overview">Machine Learning Demo room&lt;/a>.&lt;/p></description></item><item><title>Netdata Parents (Streaming and Replication)</title><link>https://www.netdata.cloud/blog/netdata-parents-streaming-replication/</link><pubDate>Fri, 30 Jun 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-parents-streaming-replication/</guid><description>&lt;h2 id="what-are-they-and-why-do-we-need-them">What are they and why do we need them?&lt;/h2>
&lt;p>A “Parent” is a Netdata Agent, like the ones we install on all our systems, but is configured as a central node that receives, stores and processes metrics data from other Netdata “Child” nodes in our infrastructure.&lt;/p>
&lt;p>Netdata Parents are flexible. You can have one big active-active cluster of Netdata Parents, or you can spread a lot of independent Parents across the infrastructure.&lt;/p></description></item><item><title>Release 1.40: Summary Tiles, Silencing &amp; ML Tweaks</title><link>https://www.netdata.cloud/blog/netdata-version-1.40/</link><pubDate>Wed, 14 Jun 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-version-1.40/</guid><description>&lt;p>Another release of the Netdata Monitoring solution is here!&lt;/p>

 &lt;div style="position: relative; padding-bottom: 56.25%; height: 0; overflow: hidden;">
 &lt;iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" allowfullscreen="allowfullscreen" loading="eager" referrerpolicy="strict-origin-when-cross-origin" src="https://www.youtube.com/embed/2VkWIZB8S30?autoplay=0&amp;amp;controls=1&amp;amp;end=0&amp;amp;loop=0&amp;amp;mute=0&amp;amp;start=0" style="position: absolute; top: 0; left: 0; width: 100%; height: 100%; border:0;" title="YouTube video">&lt;/iframe>
 &lt;/div>

&lt;!--truncate-->
&lt;ul>
&lt;li>&lt;a href="#v1400-netdata-open-source-growth">Netdata Growth&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-release-highlights">Release Highlights&lt;/a>
&lt;ul>
&lt;li>&lt;strong>&lt;a href="#v1400-visualization-summary-dashboards">Dashboard Sections&amp;rsquo; Summary Tiles&lt;/a>&lt;/strong>&lt;br/>
Added summary tiles to most sections of the fully-automated dashboards, to provide an instant view of the most important metrics for each section.&lt;/li>
&lt;li>&lt;strong>&lt;a href="#v1400-alert-notification-silencing">Silencing of Cloud Alert Notifications&lt;/a>&lt;/strong>&lt;br/>
Maintenance window coming up? Active issue being checked? Use the Alert notification silencing engine to mute your notifications.&lt;/li>
&lt;li>&lt;strong>&lt;a href="#v1400-ml-extended-training">Machine Learning - Extended Training to 24 Hours&lt;/a>&lt;/strong>&lt;br/>
Netdata now trains multiple models per metric, to learn the behavior of each metric for the last 24 hours. Trained models are persisted on disk and are loaded back on Netdata restart.&lt;/li>
&lt;li>&lt;strong>&lt;a href="#v1400-streaming">Rewritten SSL Support for the Agent&lt;/a>&lt;/strong>&lt;br/>
Netdata Agent now features a new SSL layer that allows it to reliably use SSL on all its features, including the API and Streaming.&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#v1400-alerts">Alerts and Notifications&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-visualizations">Visualizations / Charts and Dashboards&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-packaging-split">Preliminary steps to split native packages&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-acknowledgements">Acknowledgements&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-contributions">Contributions&lt;/a>
&lt;ul>
&lt;li>&lt;a href="#v1400-contributions-collectors">Collectors&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-contributions-documentation">Documentation&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-contributions-packaging">Packaging / Installation&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-contributions-streaming">Streaming&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-contributions-health">Health&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-contributions-exporting">Exporting&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-contributions-ml">ML&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-contributions-other">Other notable changes&lt;/a>&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#v1400-deprecation-notice">Deprecation notice&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-cloud-recommended-version">Cloud recommended version&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-release-meetup">Release meetup&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-support-options">Support options&lt;/a>&lt;/li>
&lt;li>&lt;a href="#v1400-running-survey">Running survey&lt;/a>&lt;/li>
&lt;/ul>
&lt;h2 id="netdata-growth-a-idv1400-netdata-open-source-growtha">Netdata Growth &lt;a id="v1400-netdata-open-source-growth">&lt;/a>&lt;/h2>
&lt;p>🚀 Our community growth is increasing steadily. ❤️ Thank you! Your love and acceptance give us the energy and passion to work harder to simplify and make monitoring easier, more effective and more fun to use.&lt;/p></description></item><item><title>How Netdata's ML-based Anomaly Detection Works</title><link>https://www.netdata.cloud/blog/how-netdatas-ml-based-anomaly-detection-works/</link><pubDate>Tue, 23 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/how-netdatas-ml-based-anomaly-detection-works/</guid><description>&lt;p>&lt;img src="../2023-05-23-how-netdatas-ml-based-anomaly-detection-works/img/img.png" alt="title image">&lt;/p>
&lt;p>How does Netdata&amp;rsquo;s &lt;a href="https://learn.netdata.cloud/docs/troubleshooting-and-machine-learning/machine-learning-ml-powered-anomaly-detection">machine learning (ML) based anomaly detection&lt;/a> actually work? Read on to find out!&lt;/p>
&lt;!--truncate-->
&lt;h2 id="design-considerations">Design considerations&lt;/h2>
&lt;p>Lets first start with some of the key design considerations and principles of Netdata&amp;rsquo;s anomaly detection (&lt;em>and some comments in parenthesis along the way&lt;/em>):&lt;/p>
&lt;ol>
&lt;li>We don&amp;rsquo;t have any labels or examples of previous anomalies. This means we are in an &lt;a href="https://en.wikipedia.org/wiki/Unsupervised_learning">unsupervised setting&lt;/a> (&lt;em>best we can try to do is learn what &amp;ldquo;normal&amp;rdquo; data looks like assuming the collected data is &amp;ldquo;mostly&amp;rdquo; normal&lt;/em>).&lt;/li>
&lt;li>Needs to be lightweight and run on the agent (&lt;em>or a parent&lt;/em>).
&lt;ul>
&lt;li>Need to be very careful of impact on CPU overhead when training and scoring (&lt;em>lots of cheap models are better than a few expensive and heavy ones&lt;/em>).&lt;/li>
&lt;li>Models themselves need to be small so as to not drastically increase the agents memory footprint (&lt;em>model objects need to be small for storage&lt;/em>).&lt;/li>
&lt;li>This has implications for the ML formulation (&lt;em>sorry - no deep learning models yet) and its implementation (we need to be surgical and optimized&lt;/em>).&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>Needs to scale for thousands of metrics and score in realtime every second as metrics are collected.
&lt;ul>
&lt;li>Typical Netdata nodes have thousands of metrics and we want to be able to score every metric every second with minimal latency overhead (&lt;em>we need to use sensible approaches to training like spreading the training cost over a wide training window&lt;/em>).&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>Needs to be able to handle a wide variety of metrics.
&lt;ul>
&lt;li>There is no single perfect model or approach for all types of metrics so we need a good all rounder that can work well enough across any and all different types of time seres metrics (&lt;em>for any given metric of course you could handcraft a better model but thats not feasible here, we need something like a &amp;ldquo;weak learners&amp;rdquo; approach of lots of generally useful models adding up to &amp;ldquo;more than the sum of their parts&amp;rdquo;&lt;/em>).&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>Needs to be written in C or C++ as that is the language of the Netdata agent (&lt;em>We are using &lt;a href="https://github.com/davisking/dlib">dlib&lt;/a> for the current implementation&lt;/em>).&lt;/li>
&lt;li>We Need to be very careful about taking big or complex dependencies if using third party libraries.
&lt;ul>
&lt;li>We want to be able to easily build and deploy Netdata on any Linux system without having to worry about installing or managing complex dependencies (&lt;em>we need to be careful of more complex algorithms that would have larger dependencies and potentially limit where Netdata can run&lt;/em>).&lt;/li>
&lt;/ul>
&lt;/li>
&lt;/ol>
&lt;p>The above considerations are important and useful to keep in mind as we explore the system in more detail.&lt;/p></description></item><item><title>Revolutionizing Ops Centers With Real-Time Monitoring</title><link>https://www.netdata.cloud/blog/revolutionizing-operations-centers-real-time-monitoring-solution/</link><pubDate>Fri, 19 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/revolutionizing-operations-centers-real-time-monitoring-solution/</guid><description>&lt;p>&lt;img src="../2023-05-19-revolutionizing-operations-centers-real-time-monitoring-solution/img/stacked-netdata.png" alt="stacked-netdata">&lt;/p>
&lt;p>In today&amp;rsquo;s fast-paced digital landscape, 24-hour operations centers play a crucial role in managing and monitoring large-scale infrastructures. These centers must be equipped with an effective monitoring solution that addresses their unique needs, enabling them to respond quickly to incidents and maintain optimal system performance. Netdata, a comprehensive monitoring solution, has been designed to meet these critical requirements with its advanced capabilities and recent enhancements.&lt;/p>
&lt;p>In this article, we will explore how Netdata&amp;rsquo;s powerful features can transform the way 24-hour operations centers monitor and manage their complex environments, leading to improved incident detection, faster troubleshooting, and better overall system performance.&lt;/p></description></item><item><title>The Future Of Infrastructure Monitoring: Scalability &amp; AI</title><link>https://www.netdata.cloud/blog/future-of-infrastructure-monitoring/</link><pubDate>Fri, 19 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/future-of-infrastructure-monitoring/</guid><description>&lt;p>In this blog post, we will explore the importance of scalability, automation, and AI in the evolving landscape of &lt;a href="https://www.netdata.cloud/blog/what-is-infrastructure-monitoring/">infrastructure monitoring&lt;/a>. We will examine how Netdata&amp;rsquo;s innovative solution aligns with these emerging trends, and how it can empower organizations to effectively manage their modern IT infrastructure.&lt;/p>
&lt;!--truncate-->
&lt;p>In today&amp;rsquo;s increasingly complex IT landscape, the need for efficient and reliable infrastructure monitoring has never been more critical. With the &lt;a href="https://medium.com/capital-one-tech/the-microservices-paradox-e55d5af2fda5">proliferation of microservices&lt;/a>, distributed systems, and cloud-native applications, managing and monitoring the performance of these rapidly evolving environments has become a significant challenge. As a result, infrastructure monitoring solutions must adapt to keep pace with these changes and deliver the insights necessary to maintain optimal performance.&lt;/p></description></item><item><title>Monitoring Multi-Cloud &amp; Hybrid-Cloud Infrastructures</title><link>https://www.netdata.cloud/blog/monitoring-multi-cloud-hybrid-cloud/</link><pubDate>Tue, 16 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/monitoring-multi-cloud-hybrid-cloud/</guid><description>&lt;p>The advent of multi-cloud and hybrid-cloud architectures has created new opportunities for organizations to leverage best-in-class features from various cloud service providers. However, these complex environments present their own unique challenges, especially when it comes to monitoring and managing performance.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="visibility">Visibility&lt;/h2>
&lt;p>The visibility challenge in multi-cloud and hybrid-cloud environments often stems from the use of disparate monitoring tools that are native to each cloud provider. While these native tools (like Amazon CloudWatch, Google Cloud Monitoring, and Azure Monitor) are excellent within their respective ecosystems, they don&amp;rsquo;t necessarily play well together when it comes to consolidating data and providing a comprehensive, unified view of your entire infrastructure.&lt;/p></description></item><item><title>Cloud Optimization: Cost, Performance &amp; Resource Strategies</title><link>https://www.netdata.cloud/blog/mastering-cloud-optimization/</link><pubDate>Sun, 14 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/mastering-cloud-optimization/</guid><description>&lt;p>Cloud optimization is the ongoing process of analyzing, configuring, and refining cloud environments to improve performance, reduce costs, and align resource usage with business needs. As cloud adoption grows, organizations must move beyond cost-cutting alone and treat optimization as a strategic practice.&lt;/p>
&lt;h2 id="cloud-optimization-strategies-to-achieve-business-goals">Cloud Optimization Strategies To Achieve Business Goals&lt;/h2>
&lt;p>Cloud optimization strategies generally focus on cost control, performance enhancement, and efficient resource utilization. These strategies range from selecting the right cloud service model (IaaS, PaaS, or SaaS), right-sizing your resources, adopting a &lt;a href="https://www.netdata.cloud/academy/what-is-cloud-management-how-to-maximize-efficiency/">multi-cloud approach&lt;/a>, automating processes, and investing in robust monitoring tools that can reliably reveal resources utilization and help you ensure that services are tailored to meet business objectives.&lt;/p></description></item><item><title>Migrating To Cloud: Key Challenges &amp; Best Practices</title><link>https://www.netdata.cloud/blog/migrating-to-cloud-key-challenges-best-practices/</link><pubDate>Sun, 14 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/migrating-to-cloud-key-challenges-best-practices/</guid><description>&lt;p>Embarking on a cloud migration journey? Grasp the obstacles and arm yourself with best practices for a smooth transition. Success lies in understanding, planning, and adapting.&lt;/p>
&lt;!--truncate-->
&lt;p>As we continue to advance further into the 21st century, businesses of all sizes are finding themselves in the midst of a digital revolution. At the heart of this transformation lies cloud migration, a process that has become a critical strategic decision for organizations aiming to remain competitive, innovative, and responsive to fluctuating market dynamics.&lt;/p></description></item><item><title>Transform Monitoring With A Machine Learning Approach</title><link>https://www.netdata.cloud/blog/transform-monitoring-ml-first-approach/</link><pubDate>Thu, 11 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/transform-monitoring-ml-first-approach/</guid><description>&lt;p>Unlocking the full potential of monitoring through ML integration, anomaly detection, and innovative scoring engines.&lt;/p>
&lt;!--truncate-->
&lt;p>Machine Learning has been making waves in various industries, but its adoption in the monitoring and observability space has been slower than expected. Many “ML” features remain gimmicky and do not provide actual real world value to users that encourages their further use.&lt;/p>
&lt;p>At Netdata, we firmly believe that ML is crucial for monitoring, and we&amp;rsquo;ve taken an ML-first approach to provide users with powerful tools and insights. In this blog post, we&amp;rsquo;ll discuss the reasons behind our belief in ML, how we&amp;rsquo;ve integrated ML into our charts and visualizations, our query engines, the scoring engine we&amp;rsquo;ve built, and how these innovations enable metrics correlations and anomaly advisor.&lt;/p></description></item><item><title>The Future of Monitoring is Automated and Opinionated</title><link>https://www.netdata.cloud/blog/the-future-of-monitoring-is-automated-and-opinionated/</link><pubDate>Tue, 09 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/the-future-of-monitoring-is-automated-and-opinionated/</guid><description>&lt;p>So, you think you monitor your infra?&lt;/p>
&lt;!-- truncate -->
&lt;p>As humanity increasingly relies on technology, &lt;a href="https://www.netdata.cloud/blog/future-of-infrastructure-monitoring/">the need for reliable and efficient infrastructure monitoring solutions has never been greater&lt;/a>.&lt;/p>
&lt;p>However, most businesses don&amp;rsquo;t take this seriously. They make poor choices that soon trap their best talent, the people who should be propelling them ahead of their competition.&lt;/p>
&lt;p>Consider this: most of the world believes that each company needs to dedicate time, talent, and money to configure and set up the monitoring of their web servers and database servers from scratch!&lt;/p></description></item><item><title>Release 1.39.0: A new era for monitoring charts.</title><link>https://www.netdata.cloud/blog/netdata-version-1.39/</link><pubDate>Mon, 08 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-version-1.39/</guid><description>&lt;p>Another release of the Netdata Monitoring solution is here!&lt;/p>

 &lt;div style="position: relative; padding-bottom: 56.25%; height: 0; overflow: hidden;">
 &lt;iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" allowfullscreen="allowfullscreen" loading="eager" referrerpolicy="strict-origin-when-cross-origin" src="https://www.youtube.com/embed/dU4GJjpeb3I?autoplay=0&amp;amp;controls=1&amp;amp;end=0&amp;amp;loop=0&amp;amp;mute=0&amp;amp;start=0" style="position: absolute; top: 0; left: 0; width: 100%; height: 100%; border:0;" title="YouTube video">&lt;/iframe>
 &lt;/div>

&lt;ul>
&lt;li>&lt;a href="#v1390-netdata-open-source-growth">Netdata open-source growth&lt;/a>&lt;/li>
&lt;li>&lt;a href="#release-highlights">Release highlights&lt;/a>
&lt;ul>
&lt;li>&lt;strong>&lt;a href="#netdata-charts-v30">Netdata Charts v3.0&lt;/a>&lt;/strong>
A new era for monitoring charts. Powerful, fast, easy to use. Instantly understand the dataset behind any chart. Slice, dice, filter and pivot the data in any way possible!&lt;/li>
&lt;li>&lt;strong>&lt;a href="#windows-support">Windows support&lt;/a>&lt;/strong>
Windows hosts are now first-class citizens. You can now enjoy out-of-the-box monitoring of over 200 metrics from your Windows systems and the services that run on them.&lt;/li>
&lt;li>&lt;a href="#virtual-nodes-and-custom-labels">Virtual nodes and custom labels&lt;/a>
You now have access to more monitoring superpowers for managing medium to large infrastructures. With custom labels and virtual hosts, you can easily organize your infrastructure and ensure that troubleshooting is more efficient.&lt;/li>
&lt;li>&lt;a href="#major-upcoming-changes">Major upcoming changes&lt;/a>
Separate packages for data collection plugins, mandatory &lt;code>zlib&lt;/code>, no upgrades of existing installs from versions prior to v1.11.&lt;/li>
&lt;li>&lt;a href="#bar-charts-for-functions">Bar charts for functions&lt;/a>&lt;/li>
&lt;li>&lt;a href="#opsgenie-notifications-for-business-plan-users">Opsgenie notifications for Business Plan users&lt;/a>
Business plan users can now seamlessly integrate Netdata with their Atlassian Opsgenie alerting and on call management system.&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#data-collection">Data Collection&lt;/a>
&lt;ul>
&lt;li>&lt;a href="#containers-and-vms-cgroups">Containers and VMs CGROUPS&lt;/a>&lt;/li>
&lt;li>&lt;a href="#docker">Docker&lt;/a>&lt;/li>
&lt;li>&lt;a href="#kubernetes">Kubernetes&lt;/a>&lt;/li>
&lt;li>&lt;a href="#kernel-tracesmetrics-ebpf">Kernel traces/metrics eBPF&lt;/a>&lt;/li>
&lt;li>&lt;a href="#disk-space-monitoring">Disk Space Monitoring&lt;/a>&lt;/li>
&lt;li>&lt;a href="#os-provided-metrics-procplugin">OS Provided Metrics proc.plugin&lt;/a>&lt;/li>
&lt;li>&lt;a href="#postgresql">PostgreSQL&lt;/a>&lt;/li>
&lt;li>&lt;a href="#dns-query">DNS Query&lt;/a>&lt;/li>
&lt;li>&lt;a href="#http-endpoint-check">HTTP endpoint check&lt;/a>&lt;/li>
&lt;li>&lt;a href="#elasticsearch-and-opensearch">Elasticsearch and OpenSearch&lt;/a>&lt;/li>
&lt;li>&lt;a href="#dnsmasq-dns-forwarder">Dnsmasq DNS Forwarder&lt;/a>&lt;/li>
&lt;li>&lt;a href="#envoy">Envoy&lt;/a>&lt;/li>
&lt;li>&lt;a href="#files-and-directories">Files and directories&lt;/a>&lt;/li>
&lt;li>&lt;a href="#rabbitmq">RabbitMQ&lt;/a>&lt;/li>
&lt;li>&lt;a href="#chartsdplugin">charts.d.plugin&lt;/a>&lt;/li>
&lt;li>&lt;a href="#anomalies">Anomalies&lt;/a>&lt;/li>
&lt;li>&lt;a href="#generic-structured-data-pandas">Generic structured data with Pandas&lt;/a>&lt;/li>
&lt;li>&lt;a href="#generic-prometheus-collector">Generic Prometheus collector&lt;/a>&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#alerts-and-notifications">Alerts and Notifications&lt;/a>
&lt;ul>
&lt;li>&lt;a href="#notifications">Notifications&lt;/a>
&lt;ul>
&lt;li>&lt;a href="#improved-email-alert-notifications">Improved email alert notifications&lt;/a>&lt;/li>
&lt;li>&lt;a href="#receive-only-notifications-for-unreachable-nodes">Receive only notifications for unreachable nodes&lt;/a>&lt;/li>
&lt;li>&lt;a href="#ntfy-agent-alert-notifications">ntfy agent alert notifications&lt;/a>&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#enhanced-real-time-alert-synchronization-on-netdata-cloud">Enhanced Real-Time Alert Synchronization on Netdata Cloud&lt;/a>&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#visualizations--charts-and-dashboards">Visualizations / Charts and Dashboards&lt;/a>
&lt;ul>
&lt;li>&lt;a href="#events-feed">Events Feed&lt;/a>&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#machine-learning">Machine Learning&lt;/a>&lt;/li>
&lt;li>&lt;a href="#installation-and-packaging">Installation and Packaging&lt;/a>
&lt;ul>
&lt;li>&lt;a href="#improved-linux-compatibility">Improved Linux compatibility&lt;/a>&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#administration">Administration&lt;/a>
&lt;ul>
&lt;li>&lt;a href="#new-way-to-retrieve-netdataconf">New way to retrieve netdata.conf&lt;/a>&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#documentation-and-demos">Documentation and Demos&lt;/a>&lt;/li>
&lt;li>&lt;a href="#deprecation-notice">Deprecation notice&lt;/a>
&lt;ul>
&lt;li>&lt;a href="#deprecated-in-this-release">Deprecated in this release&lt;/a>&lt;/li>
&lt;/ul>
&lt;/li>
&lt;li>&lt;a href="#netdata-agent-release-meetup">Netdata Agent Release Meetup&lt;/a>&lt;/li>
&lt;li>&lt;a href="#support-options">Support options&lt;/a>&lt;/li>
&lt;li>&lt;a href="#running-survey">Running survey&lt;/a>&lt;/li>
&lt;li>&lt;a href="#acknowledgements">Acknowledgements&lt;/a>&lt;/li>
&lt;/ul>
&lt;h2 id="netdata-open-source-growth">Netdata open-source growth&lt;/h2>
&lt;!-- Retrieve most of these stats from netdata/netdata/README.md badges -->
&lt;ul>
&lt;li>Over 62,000 GitHub Stars&lt;/li>
&lt;li>Over 1.5 million online nodes&lt;/li>
&lt;li>Almost 92 million sessions served&lt;/li>
&lt;li>Over 600 thousand total nodes in Netdata Cloud&lt;/li>
&lt;/ul>
&lt;h2 id="release-highlights">Release highlights&lt;/h2>
&lt;h3 id="netdata-charts-v30">Netdata Charts v3.0&lt;/h3>
&lt;p>We are excited to announce Netdata Charts v3.0 and the NIDL framework. These are currently available at Netdata Cloud. At the next Netdata release, the agent dashboard will be replaced to also use the same charts.&lt;/p></description></item><item><title>Infinite Scalability: Monitoring Without Limits</title><link>https://www.netdata.cloud/blog/netdata-inifinite-scalability/</link><pubDate>Thu, 04 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-inifinite-scalability/</guid><description>&lt;p>Scalability is crucial for monitoring systems as it ensures that they can accommodate growth, maintain performance, provide flexibility, optimize costs, enhance fault tolerance, and support informed decision-making, all of which are critical for effective infrastructure management.&lt;/p>
&lt;!--truncate-->
&lt;p>Most monitoring solutions struggle with scalability, mainly because of:&lt;/p>
&lt;ol>
&lt;li>&lt;strong>High data volume and velocity&lt;/strong>: Monitoring systems generate vast amounts of data and as the infrastructure grows, so does the volume and velocity of these data.&lt;/li>
&lt;li>&lt;strong>Resource constraints&lt;/strong>: Scalability requires efficient resource utilization, leading to bottlenecks and performance issues as the monitored environment grows.&lt;/li>
&lt;li>&lt;strong>Architectural limitations&lt;/strong>: Monitoring systems are usually designed with certain architectural constraints that limit their scalability. Most open source solutions rely on monolithic or centralized architectures that can become overwhelmed at scale.&lt;/li>
&lt;/ol>
&lt;p>For open source solutions scalability has always been a challenge, increasing their complexity significantly (check for example the scalability issues of Prometheus), while for commercial solutions it usually results in increased data collection to visualization latency and cost.&lt;/p></description></item><item><title>Monitoring Disks: Workload, Latency &amp; Saturation</title><link>https://www.netdata.cloud/blog/monitoring-disks-understanding-workload-performance-utilisation-saturation-latency/</link><pubDate>Thu, 04 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/monitoring-disks-understanding-workload-performance-utilisation-saturation-latency/</guid><description>&lt;p>&lt;img src="../2023-05-04-monitoring-disks-understanding-workload-performance-utilisation-saturation-latency/img/stacked-netdata.png" alt="stacked-netdata">&lt;/p>
&lt;p>Netdata provides a comprehensive set of charts that can help you understand the workload, performance, utilization, saturation, latency, responsiveness, and maintenance activities of your disks.
In this blog we will focus on monitoring disks as block devices, not as filesystems or mount points.&lt;/p>
&lt;!-- truncate -->
&lt;p>The &lt;code>Disks&lt;/code> section in the &lt;code>Overview&lt;/code> tab contains all the charts that are mentioned in this blog post.
&lt;img src="../2023-05-04-monitoring-disks-understanding-workload-performance-utilisation-saturation-latency/img/disks-overview.png" alt="Disks-Overview">&lt;/p>
&lt;h2 id="disk-workload-and-performance">Disk Workload and Performance&lt;/h2>
&lt;p>Netdata charts for monitoring the workload and the throughput of your disks:&lt;/p></description></item><item><title>Understanding Huge Pages</title><link>https://www.netdata.cloud/blog/understanding-huge-pages/</link><pubDate>Thu, 04 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/understanding-huge-pages/</guid><description>&lt;p>Memory-intensive applications can benefit from &lt;a href="https://www.netdata.cloud/academy/what-is-application-performance-monitoring-apm/">improved performance&lt;/a> by using huge pages, as they can reduce TLB pressure and memory fragmentation, and lower the memory management overhead overall. Developers should consider using HugeTLBfs in their mmap() and shmget() calls to take advantage of huge pages.&lt;/p>
&lt;p>Transparent Huge Pages (THP) is a Linux kernel feature that provides some of the benefits of huge pages without requiring any development effort. However, THP can cause latency in many applications. Although kernel developers are actively working to address these issues, many system administrators prefer to disable THP altogether.&lt;/p></description></item><item><title>Unlock the Secrets of Kernel Memory Usage</title><link>https://www.netdata.cloud/blog/unlock-the-secrets-of-kernel-memory-usage/</link><pubDate>Thu, 04 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/unlock-the-secrets-of-kernel-memory-usage/</guid><description>&lt;p>&lt;img src="../2023-05-04-unlock-the-secrets-of-kernel-memory-usage/img/stacked-netdata.png" alt="stacked-netdata">&lt;/p>
&lt;p>The &lt;code>mem.kernel&lt;/code> chart in Netdata provides insight into the memory usage of &lt;a href="https://www.netdata.cloud/academy/what-are-the-differences-between-bpf-and-ebpf-an-overview/">various kernel subsystems&lt;/a> and mechanisms. By understanding these dimensions and their technical details, you can monitor your system&amp;rsquo;s kernel memory usage and identify potential issues or inefficiencies. Monitoring these dimensions can help you ensure that your system is running efficiently and provide valuable insights into the performance of your kernel and memory subsystem.&lt;/p>
&lt;p>&lt;img src="../2023-05-04-unlock-the-secrets-of-kernel-memory-usage/img/mem-kernel.png" alt="mem-kernel">&lt;/p>
&lt;!-- truncate -->
&lt;h2 id="slab">Slab&lt;/h2>
&lt;p>The &lt;a href="https://en.wikipedia.org/wiki/Slab_allocation">slab allocator&lt;/a> is a memory management mechanism introduced by Jeff Bonwick in 1994 to manage &lt;a href="https://www.netdata.cloud/vsphere-monitoring/">memory allocation&lt;/a> for kernel objects. The main purpose of the slab allocator is to reduce memory fragmentation and improve the speed of memory allocation/deallocation. The slab allocator groups objects of the same size into &amp;ldquo;slabs&amp;rdquo; and caches the objects to speed up future allocations.&lt;/p></description></item><item><title>Entropy In Cryptography: Key To Security &amp; Randomness</title><link>https://www.netdata.cloud/blog/understanding-entropy-the-key-to-secure-cryptography-and-randomness/</link><pubDate>Wed, 03 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/understanding-entropy-the-key-to-secure-cryptography-and-randomness/</guid><description>&lt;p>&lt;img src="../2023-05-03-understanding-entropy-the-key-to-secure-cryptography-and-randomness/img/stacked-netdata.png" alt="stacked-netdata">&lt;/p>
&lt;p>&lt;a href="https://en.wikipedia.org/wiki/Entropy_(computing)">Entropy&lt;/a> is a measure of the randomness or unpredictability of data. In the context of cryptography, entropy is used to generate random numbers or keys that are essential for secure communication and encryption. Without a good source of entropy, cryptographic protocols can become vulnerable to attacks that exploit the predictability of the generated keys.&lt;/p>
&lt;!-- truncate -->
&lt;h2 id="what-is-entropy">What is Entropy?&lt;/h2>
&lt;p>In most operating systems, entropy is generated by collecting random events from various sources, such as hardware interrupts, mouse movements, keyboard presses, and disk activity. These events are fed into a pool of entropy, which is then used to generate random numbers when needed.&lt;/p></description></item><item><title>Context Switching &amp; Its Impact On System Performance</title><link>https://www.netdata.cloud/blog/understanding-context-switching-and-its-impact-on-system-performance/</link><pubDate>Tue, 02 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/understanding-context-switching-and-its-impact-on-system-performance/</guid><description>&lt;p>&lt;img src="../2023-05-02-understanding-context-switching-and-its-impact-on-system-performance/img/stacked-netdata.png" alt="stacked-netdata">&lt;/p>
&lt;p>Context switching is the process of switching the CPU from one process, task or thread to another. In a multitasking operating system, such as Linux, the CPU has to switch between multiple processes or threads in order to keep the system running smoothly. This is necessary because each CPU core without hyperthreading can only execute one process or thread at a time. If there are many processes or threads running simultaneously, and very few CPU cores available to handle them, the system is forced to make more context switches to balance the CPU resources among them.&lt;/p></description></item><item><title>Linux CPU Consumption, Load &amp; Pressure Explained</title><link>https://www.netdata.cloud/blog/understanding-linux-cpu-consumption-load-and-pressure-for-performance-optimisation/</link><pubDate>Tue, 02 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/understanding-linux-cpu-consumption-load-and-pressure-for-performance-optimisation/</guid><description>&lt;p>&lt;img src="../2023-05-02-understanding-linux-cpu-consumption-load-and-pressure-for-performance-optimisation/img/stacked-netdata.png" alt="stacked-netdata">&lt;/p>
&lt;p>As a system administrator, understanding how your Linux system&amp;rsquo;s CPU is being utilized is crucial for identifying bottlenecks and &lt;a href="https://www.netdata.cloud/academy/what-is-cardinality-in-databases-a-comprehensive-guide/">optimizing performance&lt;/a>. In this blog post, we&amp;rsquo;ll dive deep into the world of Linux CPU consumption, load, and pressure, and discuss how to use these metrics effectively to identify issues and improve your system&amp;rsquo;s performance.&lt;/p>
&lt;!-- truncate -->
&lt;h2 id="cpu-consumption-and-utilization">CPU Consumption and Utilization&lt;/h2>
&lt;p>CPU consumption refers to the amount of processing power being used by applications running on your system. The &lt;code>system.cpu&lt;/code> chart in Netdata represents the Total CPU utilization of your Linux system, broken down into different dimensions. Each dimension provides insight into how the CPU is being used by various tasks and processes. Here&amp;rsquo;s a brief explanation of each dimension:&lt;/p></description></item><item><title>Server Uptime Monitoring: Core Benefits For High Performance</title><link>https://www.netdata.cloud/blog/server-uptime-monitoring-why-do-we-need-it/</link><pubDate>Tue, 02 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/server-uptime-monitoring-why-do-we-need-it/</guid><description>&lt;p>&lt;img src="../2023-05-02-server-uptime-monitoring-why-do-we-need-it/img/stacked-netdata.png" alt="Server Uptime Monitoring: Core Benefits For High Performance">&lt;/p>
&lt;p>Server uptime monitoring tracks the availability and reliability of servers within your infrastructure.&lt;/p>
&lt;!-- truncate -->
&lt;h2 id="what-is-server-uptime-monitoring">What Is Server Uptime Monitoring?&lt;/h2>
&lt;p>Server uptime monitoring is the process of continuously tracking the operational status of your servers to ensure optimal performance and availability for users.&lt;/p>
&lt;p>With Netdata, you gain access to real-time, high-resolution monitoring that goes beyond basic checks, providing a detailed overview of your entire infrastructure.&lt;/p></description></item><item><title>Swap Memory: When &amp; How To Use It On Production VMs</title><link>https://www.netdata.cloud/blog/swap-memory-when-and-how-to-use-it-on-your-production-systems-or-cloud-provided-vms/</link><pubDate>Tue, 02 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/swap-memory-when-and-how-to-use-it-on-your-production-systems-or-cloud-provided-vms/</guid><description>&lt;p>&lt;img src="../2023-05-02-swap-memory-when-to-use-in-production-systems/img/stacked-netdata.png" alt="Swap Memory: Its Use On Production Systems &amp;amp; Cloud-Provided VMs">&lt;/p>
&lt;p>Swap memory, also known as virtual memory, is a space on a hard disk that is used to supplement the physical memory (RAM) of a computer. The swap space is used when the system runs out of physical memory, and it moves less frequently accessed data from RAM to the hard disk, freeing up space in RAM for more frequently accessed data. But should swap memory be enabled on production systems and &lt;a href="https://www.netdata.cloud/vsphere-monitoring/">cloud-provided virtual machines&lt;/a> (VMs)? Let&amp;rsquo;s explore the pros and cons.&lt;/p></description></item><item><title>Understanding Interrupts, Softirqs, and Softnet in Linux</title><link>https://www.netdata.cloud/blog/understanding-interrupts-softirqs-and-softnet-in-linux/</link><pubDate>Tue, 02 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/understanding-interrupts-softirqs-and-softnet-in-linux/</guid><description>&lt;p>&lt;img src="../2023-05-02-understanding-interrupts-softirqs-and-softnet-in-linux/img/stacked-netdata.png" alt="stacked-netdata">&lt;/p>
&lt;p>Interrupts, softirqs, and softnet are all critical parts of the Linux kernel that can impact system performance. In this blog post, we&amp;rsquo;ll explore their usefulness, and discuss how to monitor them using Netdata for both bare-metal servers and VMs.&lt;/p>
&lt;!-- truncate -->
&lt;h2 id="what-are-interrupts">What are Interrupts?&lt;/h2>
&lt;p>Interrupts are signals generated by hardware devices to indicate that they require attention from the CPU. Hardware devices can generate interrupts for a variety of reasons, including data transmission or reception, input/output operations, and other activities. When an interrupt is generated, the CPU stops what it is doing and handles the interrupt. Interrupts can have a significant impact on system performance, especially if there are a high number of interrupts occurring.&lt;/p></description></item><item><title>Understanding System Processes States</title><link>https://www.netdata.cloud/blog/understanding-system-processes-states/</link><pubDate>Tue, 02 May 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/understanding-system-processes-states/</guid><description>&lt;p>&lt;img src="../2023-05-02-understanding-system-processes-states/img/stacked-netdata.png" alt="stacked-netdata">&lt;/p>
&lt;p>The different states of system processes are essential to understanding how a computer system works. Each state represents a specific point in a process&amp;rsquo;s life cycle and can impact system performance and stability.&lt;/p>
&lt;!-- truncate -->
&lt;h2 id="process-states">Process States&lt;/h2>
&lt;p>Netdata&amp;rsquo;s &lt;code>system.processes_state&lt;/code> chart provides a view of these states, allowing users to monitor system performance in real-time:&lt;/p>
&lt;p>&lt;img src="../2023-05-02-understanding-system-processes-states/img/system-processes.png" alt="system-processes">&lt;/p>
&lt;ol>
&lt;li>
&lt;p>&lt;strong>Running:&lt;/strong> A process is in the Running state when it is actively using the CPU and executing instructions. This state is resource-intensive and can lead to performance issues if there are too many Running processes, causing CPU contention and system slowdowns. Processes in the Running state are prioritized using scheduling algorithms to improve system performance.&lt;/p></description></item><item><title>Why Scalable Monitoring Matters For Modern Systems</title><link>https://www.netdata.cloud/blog/why-scalable-monitoring-is-essential-for-modern-distributed-systems/</link><pubDate>Wed, 26 Apr 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/why-scalable-monitoring-is-essential-for-modern-distributed-systems/</guid><description>&lt;p>&lt;img src="../2023-04-26-why-scalable-monitoring-is-essential/img/stacked-netdata.png" alt="stacked-netdata">&lt;/p>
&lt;p>It&amp;rsquo;s becoming increasingly common to discuss the importance of scalability in monitoring solutions and how it can impact the performance and reliability of distributed systems.&lt;/p>
&lt;!-- truncate -->
&lt;p>In today&amp;rsquo;s rapidly evolving technological landscape, organizations are increasingly relying on distributed systems to power their operations. These systems consist of multiple interconnected components that work together to deliver a cohesive experience. They can span across different geographic locations, and often involve a combination of &lt;a href="https://www.netdata.cloud/product/cloud-on-premises/">on-premises, cloud&lt;/a>, and &lt;a href="https://www.netdata.cloud/solutions/technologies/docker-monitoring/">container-based environments&lt;/a>. As such, effectively managing these complex systems is critical to ensuring optimal performance, reliability, and security.&lt;/p></description></item><item><title>Netdata's AI Insights &amp; Rapid Diagnostics</title><link>https://www.netdata.cloud/blog/netdata-ai-insights-rapid-diagnostics/</link><pubDate>Wed, 19 Apr 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-ai-insights-rapid-diagnostics/</guid><description>&lt;p>Introduction to Netdata&amp;rsquo;s new visualisation providing AI Insights, supporting Rapid Diagnostics.
&lt;img src="https://user-images.githubusercontent.com/96257330/233125254-f93c9520-0a3f-4844-8d43-1f3202a5e411.png" alt="logo">&lt;/p>
&lt;!--truncate-->
&lt;h2 id="a-new-era-in-monitoring-systems-dashboards">A New Era in Monitoring Systems Dashboards&lt;/h2>
&lt;p>We&amp;rsquo;re thrilled to share an important upgrade to Netdata: &lt;strong>AI Insights &amp;amp; Rapid Diagnostics&lt;/strong>, a technology aiming to redefine what we expect from a monitoring system.&lt;/p>
&lt;h2 id="challenges-with-traditional-monitoring-dashboards">Challenges with Traditional Monitoring Dashboards&lt;/h2>
&lt;p>Traditional monitoring systems rely on a query language to help engineers create dashboards and alerts. While these languages offer power and flexibility, they come with several challenges that make monitoring and troubleshooting more complex and time-consuming:&lt;/p></description></item><item><title>Remote UNIX System Monitoring Using Net-SNMP</title><link>https://www.netdata.cloud/blog/remote-unix-monitoring-with-net-snmp/</link><pubDate>Wed, 12 Apr 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/remote-unix-monitoring-with-net-snmp/</guid><description>&lt;p>&lt;img src="../2023-04-12-remote-unix-monitoring-with-net-snmp/img/img.jpg" alt="img">&lt;/p>
&lt;p>Need to monitor a UNIX-like system, but can’t install Netdata on it? With our SNMP collector and Net-SNMP,
you can get basic system information with just a bit of relatively quick and easy configuration.&lt;/p>
&lt;!-- truncate -->
&lt;h2 id="what-is-snmp">What is SNMP?&lt;/h2>
&lt;p>The Simple Network Management Protocol, commonly known as SNMP, is a relatively lightweight protocol designed for
monitoring and configuration management for network appliances like switches, routers or gateways. However, it can also
be used for those purposes on almost any UNIX-like system thanks to the &lt;a href="http://www.net-snmp.org/">Net-SNMP project&lt;/a>.&lt;/p></description></item><item><title>Anomaly Rates in the Menu!</title><link>https://www.netdata.cloud/blog/anomaly-rates-in-the-menu/</link><pubDate>Wed, 29 Mar 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/anomaly-rates-in-the-menu/</guid><description>&lt;p>The menu (on the &lt;a href="https://learn.netdata.cloud/docs/getting-started/monitor-your-infrastructure/home-overview-and-single-node-view#overview-and-single-node-view">overview or single node tab&lt;/a>) now has an &lt;a href="https://learn.netdata.cloud/docs/troubleshooting-and-machine-learning/machine-learning-ml-powered-anomaly-detection#anomaly-rate">anomaly rate&lt;/a> button built into it that, for the entire visible window or a highlighted time range, shows the maximum chart anomaly rate within each section.&lt;/p>
&lt;p>Read on to learn more about this new feature!&lt;/p>
&lt;iframe width="560" height="315" src="https://www.youtube.com/embed/PgVh_MFHMb0?si=F2Mq6wIxHJWaHykZ" title="YouTube video player" frameborder="0" allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" referrerpolicy="strict-origin-when-cross-origin" allowfullscreen>&lt;/iframe>
&lt;h2 id="wait-what-is-an-anomaly-rate">Wait, what is an anomaly rate?&lt;/h2>
&lt;p>Netdata is the only monitoring agent that natively (for every metric, with zero config and sane defaults) produces anomaly rates in addition to just collecting raw metrics.&lt;/p></description></item><item><title>Introducing the Netdata demo space</title><link>https://www.netdata.cloud/blog/netdata-demo/</link><pubDate>Fri, 24 Mar 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-demo/</guid><description>&lt;p>&lt;img src="https://user-images.githubusercontent.com/24860547/201481889-0cf8e192-683f-4a80-9b96-4f69dd85490f.png" alt="image">&lt;/p>
&lt;p>Introducing Netdata&amp;rsquo;s Demo Space, a quick and easy way to experience monitoring environments before you set them up yourself.&lt;/p>
&lt;!--truncate-->
&lt;p>At Netdata, we are always striving to provide the best monitoring experience for our users. We understand that adopting a new monitoring solution can sometimes be challenging, especially when you&amp;rsquo;re unsure of how it will fit your specific environment. That&amp;rsquo;s why we&amp;rsquo;re excited to announce the Netdata Demo Space!&lt;/p></description></item><item><title>Upcoming Changes to Plugins in Native Packages</title><link>https://www.netdata.cloud/blog/split-plugin-packages/</link><pubDate>Wed, 15 Mar 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/split-plugin-packages/</guid><description>&lt;p>At Netdata, we’re committed to trying to make Netdata work as well as possible for our users. Sometimes though,
that means changing things in ways that aren’t exactly seamless. Such a change is coming soon for users of our
native DEB and RPM packages, and this blog post will explain what’s happening, why we’re doing it, and what
it means for our users.&lt;/p>
&lt;!-- truncate -->
&lt;h2 id="whats-changing">What’s changing?&lt;/h2>
&lt;p>Starting shortly after the v1.39.0 release of the Netdata Agent, we will be splitting most of our external
data collection plugins out to their own individual packages instead of bundling them all in the main &lt;code>netdata&lt;/code>
package. We already have this type of split for our CUPS and FreeIPMI plugins, and this new change will extend
that to also provide separate packages for the following plugins:&lt;/p></description></item><item><title>Windows Server Monitoring Improvements</title><link>https://www.netdata.cloud/blog/windows-monitoring-improvements/</link><pubDate>Mon, 13 Mar 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/windows-monitoring-improvements/</guid><description>&lt;p>&lt;a href="https://www.netdata.cloud/blog/web-servers-and-their-performance/">Monitor your Windows server and applications&lt;/a> running on it with Netdata - simple, powerful and free.&lt;/p>
&lt;!--truncate-->
&lt;p>Hey Netdata community,&lt;/p>
&lt;p>We have some exciting news for you: we’re launching our new and updated &lt;a href="https://learn.netdata.cloud/docs/data-collection/monitor-anything/System%20Metrics/Windows-machines">Windows collectors&lt;/a> with the goal of making the &lt;a href="https://www.netdata.cloud/windows-monitoring/">Windows monitoring experience&lt;/a> as seamless as possible 🎉&lt;/p>
&lt;p>We know that Windows monitoring has been a long time ask from many of you, and we’ve been working hard to make it easier than ever to monitor your Windows metrics with Netdata.&lt;/p></description></item><item><title>Anomaly detection on Prometheus metrics</title><link>https://www.netdata.cloud/blog/anomaly-detection-on-prometheus-metrics/</link><pubDate>Wed, 01 Mar 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/anomaly-detection-on-prometheus-metrics/</guid><description>&lt;p>&lt;img src="../2023-03-01-anomaly-detection-on-prometheus-metrics/img/img.png" alt="img">&lt;/p>
&lt;p>We have recently extended the native machine learning (ML) based anomaly detection &lt;a href="https://learn.netdata.cloud/guides/monitor/anomaly-detection">capabilities&lt;/a> of Netdata to &lt;a href="https://github.com/netdata/netdata/issues/14218">support all metrics&lt;/a>, regardless on their collection frequency (&lt;code>update every&lt;/code>).&lt;/p>
&lt;p>Previously only metrics collected every second were supported, but now Netdata can run anomaly detection out of the box with zero config on metrics with any collection frequency.&lt;/p>
&lt;p>This post will illustrate an example of what this means using &lt;a href="https://prometheus.io/">Prometheus&lt;/a> metrics (via the &lt;a href="https://learn.netdata.cloud/docs/agent/collectors/go.d.plugin/modules/prometheus#gsc.tab=0">Netdata Prometheus collector&lt;/a>) since they typically have a default collection frequency of 10 seconds.&lt;/p></description></item><item><title>Monitor any SQL metrics with Netdata (and Pandas ❤️)</title><link>https://www.netdata.cloud/blog/monitor-any-sql-metrics-with-netdata/</link><pubDate>Wed, 22 Feb 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/monitor-any-sql-metrics-with-netdata/</guid><description>&lt;p>&lt;img src="../2023-02-22-monitor-any-sql-metrics-with-netdata/img/img.png" alt="img">&lt;/p>
&lt;p>We recently got this great feedback from a dear user in our &lt;a href="https://discord.com/channels/847502280503590932/1075370683393118278/1075723915265069106">Discord&lt;/a>:&lt;/p>
&lt;blockquote>
&lt;p>I would really like to use Netdata to monitor custom internal metrics that come from SQL, not a fan of having 10 diff systems doing essentially the same thing as is, Netdata is pretty much all there in that regard, just needs a few extra features.&lt;/p>
&lt;/blockquote>
&lt;p>This is great and exactly what we want, a clear problem or improvement we could make to help make that users monitoring life a little easier.&lt;/p></description></item><item><title>Introducing Netdata Functions (↑ Top)</title><link>https://www.netdata.cloud/blog/netdata-functions/</link><pubDate>Wed, 15 Feb 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-functions/</guid><description>&lt;p>Netdata is committed to making it simpler and easier for everyone to monitor and troubleshoot their infrastructure. With that goal in mind, we&amp;rsquo;re excited to announce the launch of our new &amp;ldquo;Functions&amp;rdquo; feature (↑ Top), which allows Netdata Agent collectors to expose &amp;ldquo;functions&amp;rdquo; that can be executed in run-time and on-demand.&lt;/p>
&lt;!--truncate-->
&lt;h3 id="what-are-netdata-functions">What are Netdata functions?&lt;/h3>
&lt;p>Netdata has always been synonymous with real time monitoring and automated dashboards, with the recent introduction of &amp;ldquo;functions&amp;rdquo;, there&amp;rsquo;s now a new way for users to troubleshoot their infrastructure. &lt;strong>A function, in the context of Netdata, is a routine or script that can be invoked to run on a node and retrieve useful information, which is then displayed in the Netdata cloud dashboard.&lt;/strong>&lt;/p></description></item><item><title>Introducing Netdata Paid Subscriptions</title><link>https://www.netdata.cloud/blog/introducing-netdata-paid-subscriptions/</link><pubDate>Fri, 10 Feb 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/introducing-netdata-paid-subscriptions/</guid><description>&lt;p>All Netdata functionality is and will be available for free forever in the Community Plan. Paid tiers include features targeted for businesses and users who would need to customise their monitoring solution with different levels of user access, extra notification mechanisms, customer support and more.&lt;/p>
&lt;!--truncate-->
&lt;p>&lt;strong>Hello Netdata community&lt;/strong>,&lt;/p>
&lt;p>We are excited to announce that we are introducing new &lt;strong>Paid Subscriptions&lt;/strong> to Netdata as of Wednesday, 22nd of February 2023.&lt;/p></description></item><item><title>Release 1.38: Dramatic Performance &amp; Stability Gains</title><link>https://www.netdata.cloud/blog/netdata-version-1.38/</link><pubDate>Mon, 06 Feb 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-version-1.38/</guid><description>&lt;p>Another release of the Netdata Monitoring solution is here!&lt;/p>

 &lt;div style="position: relative; padding-bottom: 56.25%; height: 0; overflow: hidden;">
 &lt;iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" allowfullscreen="allowfullscreen" loading="eager" referrerpolicy="strict-origin-when-cross-origin" src="https://www.youtube.com/embed/2EjKicsRYxw?autoplay=0&amp;amp;controls=1&amp;amp;end=0&amp;amp;loop=0&amp;amp;mute=0&amp;amp;start=0" style="position: absolute; top: 0; left: 0; width: 100%; height: 100%; border:0;" title="YouTube video">&lt;/iframe>
 &lt;/div>

&lt;ul>
&lt;li>&lt;a href="#v1380-release-highlights">Release Highlights&lt;/a>
&lt;ul>
&lt;li>
&lt;p>&lt;strong>&lt;a href="#v1380-dbenginev2">DBENGINE v2&lt;/a>&lt;/strong>
The new open-source database engine for Netdata Agents, offering huge performance, scalability and stability improvements, with a fraction of memory footprint!&lt;/p>
&lt;/li>
&lt;li>
&lt;p>&lt;strong>&lt;a href="#v1380-functions">FUNCTION: Processes&lt;/a>&lt;/strong>
Netdata beyond metrics! We added the ability for &lt;strong>runtime functions&lt;/strong>, that can be implemented by any data collection plugin, to offer unlimited visibility to anything, even not-metrics, that can be valuable while troubleshooting.&lt;/p></description></item><item><title>Extending Netdata's anomaly detection training window</title><link>https://www.netdata.cloud/blog/extending-anomaly-detection-training-window/</link><pubDate>Thu, 02 Feb 2023 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/extending-anomaly-detection-training-window/</guid><description>&lt;p>We have been busy at work under the hood of the Netdata agent to introduce new capabilities that let you extend the &amp;ldquo;training window&amp;rdquo; used by Netdata&amp;rsquo;s &lt;a href="https://learn.netdata.cloud/docs/nightly/setup/configure-machine-learning-ml-powered-anomaly-detection">native anomaly detection capabilities&lt;/a>.&lt;/p>
&lt;p>This blog post will discuss one of these improvements to help you reduce &amp;ldquo;&lt;a href="https://en.wikipedia.org/wiki/False_positives_and_false_negatives#False_positive_error">false positives&lt;/a>&amp;rdquo; by essentially extending the training window by using the new (beautifully named) &lt;code>number of models per dimension&lt;/code> configuration parameter.&lt;/p>
&lt;h2 id="background">Background&lt;/h2>
&lt;p>One of the most important considerations of our native anomaly detection capabilities is the overhead of running the training and scoring computations required to train thousands of models (one per metric) and produce &lt;a href="https://learn.netdata.cloud/docs/nightly/setup/configure-machine-learning-ml-powered-anomaly-detection#anomaly-bit">anomaly bits&lt;/a> every second based on those trained models.&lt;/p></description></item><item><title>Release 1.37.1: Patch Release For Security Issues</title><link>https://www.netdata.cloud/blog/netdata-version-1.37.1/</link><pubDate>Mon, 05 Dec 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-version-1.37.1/</guid><description>&lt;p>Netdata v1.37.1 is a patch release to address issues discovered since v1.37.0. Refer to the &lt;a href="https://github.com/netdata/netdata/releases/tag/v1.37.0">v.1.37.0 release notes&lt;/a> for the full scope of that release.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="release-v1371">Release v1.37.1&lt;/h2>
&lt;p>Netdata v1.37.1 is a patch release to address issues discovered since v1.37.0. Refer to the &lt;a href="https://github.com/netdata/netdata/releases/tag/v1.37.0">v.1.37.0 release notes&lt;/a> for the full scope of that release.&lt;/p>
&lt;p>The v1.37.1 patch release fixes the following issues:&lt;/p>
&lt;ul>
&lt;li>Parent agent crash when many children instances (re)connect at the same time, causing simultaneous SSL re-initialization (&lt;a href="https://github.com/netdata/netdata/pull/14076">PR #14076&lt;/a>).&lt;/li>
&lt;li>Agent crash during dbengine database file rotation while a page is being read while being deleted (&lt;a href="https://github.com/netdata/netdata/pull/14081">PR #14081&lt;/a>).&lt;/li>
&lt;li>Agent crash on metrics page alignment when metrics were stopped being collected for a long time and then started again (&lt;a href="https://github.com/netdata/netdata/pull/14086">PR #14086&lt;/a>).&lt;/li>
&lt;li>Broken Fedora native packages (&lt;a href="https://github.com/netdata/netdata/pull/14082">PR #14082&lt;/a>).&lt;/li>
&lt;li>Fix dbengine backfilling statistics (&lt;a href="https://github.com/netdata/netdata/pull/14074">PR #14074&lt;/a>).&lt;/li>
&lt;/ul>
&lt;p>In addition, the release contains the following optimizations and improvements:&lt;/p></description></item><item><title>Release 1.37: Infinite Scalability &amp; Database Tiering</title><link>https://www.netdata.cloud/blog/netdata-version-1.37/</link><pubDate>Wed, 30 Nov 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-version-1.37/</guid><description>&lt;p>Another release of the Netdata Monitoring solution is here!&lt;/p>
&lt;p>We focused on these key areas:&lt;/p>
&lt;p>Infinite scalability of the Netdata Ecosystem&lt;/p>
&lt;p>Default Database Tiering, offering months of data retention for typical Netdata Agent installations with default settings and years of data retention for dedicated Netdata Parents.&lt;/p>
&lt;p>Overview Dashboards at Netdata Cloud got a ton of improvements to allow slicing and dicing of data directly on the UI and overcome the limitations of the web technology when thousands of charts are presented on one page.&lt;/p></description></item><item><title>Monitor &amp; Troubleshoot ISP Performance With Netdata</title><link>https://www.netdata.cloud/blog/speedtest-monitoring/</link><pubDate>Mon, 28 Nov 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/speedtest-monitoring/</guid><description>&lt;p>Find out how to monitor your Internet speed and quality and how well your ISP is performing.&lt;/p>
&lt;p>&lt;img src="https://user-images.githubusercontent.com/24860547/204470316-4682e442-6df1-4c77-b1e4-fdd96dd404f0.jpg" alt="logo">&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-factors-affect-my-internet-speed">What Factors Affect My Internet Speed?&lt;/h2>
&lt;p>Several factors can influence your internet speed, ranging from your ISP&amp;rsquo;s infrastructure to your home setup. Here&amp;rsquo;s a breakdown:&lt;/p>
&lt;h3 id="1-isp-plan--bandwidth">1. ISP Plan &amp;amp; Bandwidth&lt;/h3>
&lt;p>The speed you experience depends on the plan you choose from your ISP. Higher-tier plans offer more bandwidth, which means faster speeds for downloading, streaming, and gaming.&lt;/p></description></item><item><title>How to monitor node reboots?</title><link>https://www.netdata.cloud/blog/monitoring-node-reboots/</link><pubDate>Thu, 17 Nov 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/monitoring-node-reboots/</guid><description>&lt;p>Monitoring the health and status of nodes and servers is a critical part of effective infrastructure monitoring.&lt;/p>
&lt;p>&lt;img src="https://user-images.githubusercontent.com/96257330/202475049-22838a0b-73b1-485b-8416-5fd49d6ccb53.png" alt="logo">&lt;/p>
&lt;!--truncate-->
&lt;h2 id="how-to-monitor-node-reboots">How to monitor node reboots?&lt;/h2>
&lt;p>One of the most critical tasks of monitoring an infrastructure is to check the health of its servers/nodes. In most cases, this results in setting up a &amp;ldquo;Hardware manager&amp;rdquo; from the hardware vendor delivering these servers or setting up an SNMP (or similar) agent to continuously monitor the availability of the server and report when there is a reboot / failure.&lt;/p></description></item><item><title>How To Mute Alerts During Maintenance Windows</title><link>https://www.netdata.cloud/blog/mute-alerts/</link><pubDate>Thu, 03 Nov 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/mute-alerts/</guid><description>&lt;p>The health management APIs in Netdata allows teams to eliminate unnecessary alerting during scheduled maintenance, testing, auto scaling events, and instance reboots.&lt;/p>
&lt;!--truncate-->
&lt;p>For all SREs, it is absolutely crucial to filter out expected events during maintenance windows and quickly pinpoint critical issues in your infrastructure. Every minute is crucial while dealing with troubleshooting issues and any distractions that may hijack the troubleshooting process should be subdued.
The health &lt;a href="https://www.netdata.cloud/blog/iot-monitoring-challenges/">management APIs&lt;/a> in Netdata allows teams to eliminate unnecessary alerting during scheduled maintenance, testing, &lt;a href="https://www.netdata.cloud/blog/server-uptime-monitoring-why-do-we-need-it/">auto scaling events&lt;/a>, and instance reboots.&lt;/p></description></item><item><title>Monitor indoor air quality with Airthings and Netdata</title><link>https://www.netdata.cloud/blog/airquality/</link><pubDate>Wed, 02 Nov 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/airquality/</guid><description>&lt;p>Monitoring indoor air quality with &lt;a href="https://www.airthings.com/">Airthings&lt;/a> and Netdata. Understanding and measuring common contaminants and pollutants reduces your risk of air quality health concerns.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="indoor-air-quality-and-what-to-monitor">Indoor air quality and what to monitor&lt;/h2>
&lt;p>Indoor air quality can be a crucial influence on your health, wellbeing and productivity.&lt;/p>
&lt;p>Understanding and measuring common contaminants and pollutants is the first step towards reducing your risk of air quality health concerns.&lt;/p>
&lt;p>Airthings is a company that makes great air quality sensors that measure a wide variety of different variables including:&lt;/p></description></item><item><title>Monitor KSM performance with Netdata</title><link>https://www.netdata.cloud/blog/ksm/</link><pubDate>Tue, 01 Nov 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/ksm/</guid><description>&lt;p>Monitoring KSM (Kernel Same-page Merging) performance at deduping memory shared across VMs.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="kernel-same-page-merging-ksm">Kernel Same-page Merging (KSM)&lt;/h2>
&lt;p>Linux kernels store memory in &lt;strong>pages&lt;/strong> which are moved in and out of memory as a single block. On most Linux architectures pages are 4096 bytes. &lt;strong>KSM&lt;/strong> (Kernel Same-page Merging) is a kernel feature that scans memory looking for pages with identical content, and then de-duplicates them. The most common use-case where such duplicate pages occur is on hosts running multiple virtual machines (VMs).&lt;/p></description></item><item><title>Monitoring &amp; troubleshooting Cassandra with Netdata</title><link>https://www.netdata.cloud/blog/cassandra-monitoring-part2/</link><pubDate>Sat, 29 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/cassandra-monitoring-part2/</guid><description>&lt;p>How to monitor and troubleshoot Cassandra with Netdata.&lt;/p>
&lt;p>&lt;img src="https://user-images.githubusercontent.com/24860547/198524087-37dda416-a9a9-4c55-b379-0f46e990f83f.png" alt="logo">&lt;/p>
&lt;!--truncate-->
&lt;p>&lt;em>&lt;strong>Note&lt;/strong>: This post is the second part of a Cassandra monitoring series. Be sure to read our first entry &lt;a href="https://blog.netdata.cloud/cassandra-monitoring-part1">here&lt;/a>.&lt;/em>&lt;/p>
&lt;h2 id="monitoring-cassandra-with-netdata">Monitoring Cassandra with Netdata&lt;/h2>
&lt;p>Netdata’s &lt;a href="https://learn.netdata.cloud/docs/agent/collectors/go.d.plugin/modules/cassandra">Cassandra collector documentation&lt;/a> explains how to set it up to collect metrics automatically.&lt;/p>
&lt;p>Once you have followed the instructions in the docs and have installed and configured Netdata on the Cassandra cluster you are ready to start monitoring and troubleshooting. Check out the &lt;a href="https://app.netdata.cloud/spaces/netdata-demo/rooms/cassandra/overview">Cassandra demo room&lt;/a> to interact with the charts, metrics and other functionality described here.&lt;/p></description></item><item><title>How to monitor and fix Database bloats in PostgreSQL?</title><link>https://www.netdata.cloud/blog/postgresql-database-bloat/</link><pubDate>Fri, 28 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/postgresql-database-bloat/</guid><description>&lt;p>Database bloat is disk space that was used by a table or index and is available for reuse by the database but has not been reclaimed. Bloat is created when deleting or updating tables and indexes. Here&amp;rsquo;s how to deal with it!&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-is-database-bloat">What is Database bloat?&lt;/h2>
&lt;p>Database bloat is disk space that was used by a table or index and is available for reuse by the database but has not been reclaimed. Bloat is created when deleting or updating tables and indexes.&lt;/p></description></item><item><title>Cassandra Monitoring: Key Metrics &amp; Best Practices</title><link>https://www.netdata.cloud/blog/cassandra-monitoring-part1/</link><pubDate>Thu, 27 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/cassandra-monitoring-part1/</guid><description>&lt;p>What are the important Cassandra metrics to monitor and how to monitor them.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-is-cassandra--why-use-it">What Is Cassandra &amp;amp; Why Use It&lt;/h2>
&lt;p>Cassandra is an open-source, distributed, wide-column NoSQL database management system written in Java. Cassandra was originally developed by &lt;a href="https://twitter.com/hedvigeng">Avinash Lakshmanan&lt;/a> and &lt;a href="https://twitter.com/pmalik">Prashant Malik&lt;/a> at Facebook and then released as open source, eventually becoming part of the &lt;a href="https://www.netdata.cloud/apache-monitoring/">Apache&lt;/a> project.&lt;/p>
&lt;p>&lt;a href="https://www.netdata.cloud/integrations/data-collection/databases/cassandra/">Cassandra is a NoSQL database&lt;/a> - NoSQL (also known as &amp;ldquo;not only SQL&amp;rdquo;) databases do not require data to be stored in tabular format. They provide flexible schemas and scale easily with large amounts of data and high user loads.&lt;/p></description></item><item><title>How to find out which application is causing server load</title><link>https://www.netdata.cloud/blog/server-load/</link><pubDate>Wed, 26 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/server-load/</guid><description>&lt;p>We often hear the term load used to describe the state of a server or a device, but we&amp;rsquo;re here to tell you what it means, precisely, and how to monitor it.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-is-server-load">What is server load?&lt;/h2>
&lt;p>We often hear the term &amp;ldquo;load&amp;rdquo; used to describe the state of a server or a &lt;a href="https://www.netdata.cloud/blog/iot-monitoring-challenges/">device&lt;/a>. But what does it really mean?&lt;/p>
&lt;p>System load is a measure of the amount of computational work that a system performs. An overloaded system, by definition, isn&amp;rsquo;t able to complete all its
tasks per schedule - this affects the performance and productivity of the system. And while &amp;ldquo;load&amp;rdquo; often gets conflated with CPU usage there&amp;rsquo;s a lot more to it.&lt;/p></description></item><item><title>How to monitor the disk usage on your infrastructure</title><link>https://www.netdata.cloud/blog/disk-usage/</link><pubDate>Tue, 25 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/disk-usage/</guid><description>&lt;p>The most important part of disk usage monitoring is to check the utilization of each filesystem and each mount point which can reveal existing or impending issues with the storage space on your infrastructure.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="what-does-disk-usage-du-mean">What Does Disk Usage (DU) Mean?&lt;/h2>
&lt;p>Disk usage (DU) refers to the portion or percentage of computer storage that is currently in use. It contrasts with disk space or &lt;a href="https://www.netdata.cloud/blog/server-uptime-monitoring-why-do-we-need-it/">capacity&lt;/a>,
which is the total amount of space that a given disk is capable of storing. Disk usage is a crucial metric to any computing system,
as it gives the user the information needed not only for storage, but also software requirements and overall operation. Although it usually
refers to a computer’s hard disk, it may also refer to external storage, such as a USB drive or compact disc (CD).&lt;/p></description></item><item><title>7 types of Redis latency and how to fix it</title><link>https://www.netdata.cloud/blog/7-types-of-redis-latency/</link><pubDate>Mon, 24 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/7-types-of-redis-latency/</guid><description>&lt;p>Redis is designed to be fast. In most cases, it is. However, there are times when Redis may be slow, due to network issues, disk latency, or other factors. When this happens, it is important to be able to &lt;a href="https://www.netdata.cloud/blog/server-uptime-monitoring-why-do-we-need-it/">detect the slow down&lt;/a> and investigate the cause of Redis latency.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="redis--latency">Redis &amp;amp; latency&lt;/h2>
&lt;p>Latency is the maximum delay between the time a client issues a command and the time the reply to the command is received by the client. Redis has strict requirements on average and worst case latency.&lt;/p></description></item><item><title>How to monitor systemd service liveness</title><link>https://www.netdata.cloud/blog/systemd-service-liveness/</link><pubDate>Fri, 21 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/systemd-service-liveness/</guid><description>&lt;p>The life of a sysadmin or SRE is often difficult, but occasionally very simple things can make a huge difference. Basic monitoring of your systemd services is one of those simple things, which we sometimes overlook. The simplest question one would want to know is if the thing that’s supposed to be running is actually running at all. If you use systemd services, you can guarantee an answer to that question within minutes using Netdata.&lt;/p></description></item><item><title>Web Servers Monitoring: Key Metrics &amp; Strategies</title><link>https://www.netdata.cloud/blog/web-servers-and-their-performance/</link><pubDate>Thu, 20 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/web-servers-and-their-performance/</guid><description>&lt;!--truncate-->
&lt;h2 id="the-importance-of-monitoring-web-servers">The Importance Of Monitoring Web Servers&lt;/h2>
&lt;p>Web servers are among the most important components in modern &lt;a href="https://www.netdata.cloud/blog/future-of-infrastructure-monitoring/">IT infrastructures&lt;/a>. They host the websites, web services, and &lt;a href="https://www.netdata.cloud/dotnet-monitoring/">web applications&lt;/a> that we use on a daily basis. Social networking, media streaming, software as a service (SaaS), and other activities wouldn’t be possible without the use of web servers. And with the advent of cloud computing and the movement of more services online, &lt;a href="https://www.netdata.cloud/blog/server-uptime-monitoring-why-do-we-need-it/">web servers and their monitoring are only becoming more important&lt;/a>. Given the extensive usage of Web servers, Sysadmins and SREs should monitor web servers as a key aspect for performance. &lt;/p></description></item><item><title>Using Pandas In Python: Data Analysis &amp; Performance Insights</title><link>https://www.netdata.cloud/blog/pandas-python/</link><pubDate>Wed, 19 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/pandas-python/</guid><description>&lt;p>Netdata just got a &lt;a href="https://learn.netdata.cloud/docs/agent/collectors/python.d.plugin/pandas" target="_blank" rel="noopener">Pandas collector&lt;/a>.&lt;/p>
&lt;!--truncate-->
&lt;p>Pandas is a de-facto standard in reading and processing most types of structured data in Python so if you have some csv/json/xml data, either locally or via some HTTP endpoint, containing metrics you&amp;rsquo;d like to monitor, chances are you can now easily do this by leveraging the Pandas collector without having to develop your own custom collector as you might have in the past.&lt;/p></description></item><item><title>How to monitor HTTP endpoints</title><link>https://www.netdata.cloud/blog/http-endpoints/</link><pubDate>Mon, 17 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/http-endpoints/</guid><description>&lt;p>The &lt;a href="https://en.wikipedia.org/wiki/Hypertext_Transfer_Protocol">HTTP protocol&lt;/a> has become the de facto standard application layer protocol of the internet. From publicly available web sites and APIs to “inter-process” communications in REST based microservice architectures or large &lt;a href="https://en.wikipedia.org/wiki/Service-oriented_architecture">Service Oriented Architectures&lt;/a> based on &lt;a href="https://en.wikipedia.org/wiki/SOAP">SOAP&lt;/a>, you find HTTP being used again and again, due to its simplicity and our familiarity with it. How many protocols can you name that have &lt;a href="https://imgur.com/gallery/4KqWq">memes&lt;/a> for their status codes? Of course, such a popular protocol has endless pages written about how to properly monitor the services that rely on it, with many options specific to every use case.&lt;!--truncate--> What you will learn here is how to get your basics done in monitoring HTTP endpoints, so you can be up and running in a few minutes, monitoring all HTTP services in your &lt;a href="https://www.netdata.cloud/blog/future-of-infrastructure-monitoring/">entire infrastructure&lt;/a>. &lt;/p></description></item><item><title>How to monitor DNS query response time</title><link>https://www.netdata.cloud/blog/dns-query-response-time/</link><pubDate>Wed, 12 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/dns-query-response-time/</guid><description>&lt;p>DNS (Domain Name System) servers translate standard language web addresses to their actual IP addresses for network access.&lt;/p>
&lt;p>&lt;a href="https://xiaolishen.medium.com/the-dns-lookup-journey-240e9a5d345c">&lt;strong>DNS Lookup Journey&lt;/strong>&lt;/a>
&lt;img src="../wp-archive/uploads/2022/10/DNS-1.png" alt="">&lt;/p>
&lt;!--truncate-->
&lt;p>DNS response time is the time it takes a Domain Name Server to receive the request for a domain name’s IP address, process it, and return the IP address to the browser or application requesting it. When it comes to DNS response times, the lower the better, and generally values less than 100ms are considered to be in the acceptable range (depending on the application).&lt;/p></description></item><item><title>Why is data replication important?</title><link>https://www.netdata.cloud/blog/why-is-data-replication-important/</link><pubDate>Wed, 12 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/why-is-data-replication-important/</guid><description>&lt;p>High availability. This is what every monitoring tool needs to ensure that you never compromise on IT infrastructure visibility.&lt;!--truncate--> On top of high availability, do you really want to enable all available features on your production system? It is important for the monitoring tool to have a low footprint on your CPU consumption and memory usage. Let’s dive deeper into the recommended way of configuring Netdata to ensure high availability and a low resource footprint through data replication.&lt;/p></description></item><item><title>How to monitor host reachability</title><link>https://www.netdata.cloud/blog/how-to-monitor-host-reachability/</link><pubDate>Mon, 10 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/how-to-monitor-host-reachability/</guid><description>&lt;p>Most sysadmins and developers have at some point used a few of the popular &lt;a href="https://www.tecmint.com/linux-networking-commands" target="_blank" rel="noopener">Linux networking commands&lt;/a> or their Windows equivalents to answer the common questions of &lt;a href="https://www.netdata.cloud/blog/web-servers-and-their-performance/">host reachability&lt;/a> - that is, whether a host or service is reachable and how fast it responds.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="common-approaches-to-reachability">Common approaches to reachability&lt;/h2>
&lt;p>One of the simplest, common checks, is to simply &lt;code>ping&lt;/code> a host to verify that it’s reachable from where you issue the command, and to see the total time it takes for the host to receive your request.&lt;/p></description></item><item><title>Introducing the Netdata Source Plugin for Grafana</title><link>https://www.netdata.cloud/blog/introducing-netdata-source-plugin-for-grafana/</link><pubDate>Fri, 07 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/introducing-netdata-source-plugin-for-grafana/</guid><description>&lt;p>&lt;img src="../wp-archive/uploads/2022/09/postgresql_dash-600x354.png" alt="sample-dashboard">&lt;/p>
&lt;p>The open-source community is about to benefit greatly from Netdata&amp;rsquo;s new Grafana data source plugin, which makes use of a powerful data collection engine.&lt;/p>
&lt;!--truncate-->
&lt;p>This new plugin maximizes the troubleshooting capabilities of Netdata in Grafana, making them more widely available. Some of the key capabilities provided to you with this plugin include the following:&lt;/p>
&lt;ul>
 	&lt;li>Real-time monitoring with single-second granularity.&lt;/li>
 	&lt;li>Installation and out-of-the-box integrations available in seconds from one line of code.&lt;/li>
 	&lt;li>2,000+ metrics from across your whole Infrastructure, with insightful metadata associated with them.&lt;/li>
 	&lt;li>Access to our fresh ML metrics (anomaly rates) - exposing our ML capabilities at the edge!&lt;/li>
&lt;/ul>
&lt;h2 id="why-did-we-decide-to-do-it">Why did we decide to do it?&lt;/h2>
&lt;p>We are huge fans of Open-Source culture. Open-source is deeply rooted in Netdata&amp;rsquo;s DNA. Because of this, at Netdata, we don’t really buy into the “single pane of glass” or “observability platform” buzzwords. The reality is that things are just more complicated than that in real life.&lt;/p></description></item><item><title>How to filter metrics by label?</title><link>https://www.netdata.cloud/blog/how-to-filter-metrics-by-label/</link><pubDate>Thu, 06 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/how-to-filter-metrics-by-label/</guid><description>&lt;p>It is sometimes easy to get lost in the mountain of metrics and infinite number of dimensions when working with an infrastructure monitoring tool. Being able to filter metrics by label and visualize only what is relevant to the current scope of monitoring &amp;amp;troubleshooting, becomes absolutely crucial to the success of SREs, Sysadmins and DevOps professionals.&lt;/p>
&lt;!--truncate-->
&lt;p>The Netdata &lt;a href="https://staging1--netdata-docusaurus.netlify.app/docs/getting-started/netdata-in-a-pane">chart label filtering feature&lt;/a> supports grouping by and filtering each chart based on labels (key/value pairs) applicable to the context and provides fine-grain capability on slicing the data / metrics.&lt;/p></description></item><item><title>Missing indexes in PostgreSQL? How to quickly identify it</title><link>https://www.netdata.cloud/blog/missing-indexes-in-postgresql/</link><pubDate>Wed, 05 Oct 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/missing-indexes-in-postgresql/</guid><description>&lt;p>While working on improving the &lt;a href="https://netdata.cloud/postgresql-monitoring/">Netdata PostgreSQL collector&lt;/a>, we were monitoring our production PostgreSQL instance and something caught our attention immediately. The rows fetched ratio seemed really, really low for one particular database&amp;hellip; there were missing indexes in PostgreSQL!&lt;/p>
&lt;!--truncate-->
&lt;p>&lt;b>Rows fetched ratio&lt;/b> is the percentage of rows that contain data needed to execute the query (rows fetched), out of the total number of rows scanned (rows returned). A low value indicates that the database is performing extra work by scanning a large number of rows that aren’t required to process the query.&lt;/p></description></item><item><title>Data Collection Strategies For Infrastructure</title><link>https://www.netdata.cloud/blog/data-collection-strategies/</link><pubDate>Tue, 06 Sep 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/data-collection-strategies/</guid><description>&lt;!--truncate-->
&lt;p>Monitoring and troubleshooting; unfortunately, these terms are still used interchangeably, which can lead to misunderstandings about data collection strategies.&lt;/p>
&lt;p>In this article we aim to clarify some important definitions, processes, and common data collection strategies for monitoring solutions. We will specify the limitations of the described strategies, as well as key benefits which can potentially be also used for troubleshooting needs.&lt;/p>
&lt;p>&lt;strong>IT infrastructure monitoring&lt;/strong> is a business process of collecting and analyzing data over a period of time to improve business results.&lt;/p></description></item><item><title>How Netdata’s Machine Learning works</title><link>https://www.netdata.cloud/blog/how-netdatas-machine-learning-works/</link><pubDate>Thu, 01 Sep 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/how-netdatas-machine-learning-works/</guid><description>&lt;p>Following on from the &lt;a href="https://www.netdata.cloud/blog/introducing-anomaly-advisor-unsupervised-anomaly-detection-in-netdata" target="_blank" rel="noopener">recent launch&lt;/a> of our &lt;a href="https://learn.netdata.cloud/docs/cloud/insights/anomaly-advisor" target="_blank" rel="noopener">Anomaly Advisor&lt;/a> feature, and in keeping with &lt;a href="https://www.netdata.cloud/blog/our-approach-to-machine-learning/" target="_blank" rel="noopener">our approach to machine learning&lt;/a>, &lt;a href="https://github.com/netdata/netdata/blob/master/ml/notebooks/netdata_anomaly_detection_deepdive.ipynb" target="_blank" rel="noopener">here&lt;/a> is a detailed Python notebook outlining exactly how the machine learning powering the Anomaly Advisor actually works under the hood.&lt;/p>
&lt;!--truncate-->
&lt;p>Or if you&amp;rsquo;d rather watch a video walkthrough of the notebook then check out below.&lt;/p>
&lt;iframe width="560" height="315" src="https://www.youtube.com/embed/L1xleckyuDQ?si=rptYzWE-eLlhSL9x" title="YouTube video player" frameborder="0" allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" referrerpolicy="strict-origin-when-cross-origin" allowfullscreen>&lt;/iframe>
&lt;p>Try it for yourself, &lt;a href="https://learn.netdata.cloud/docs/cloud/get-started" target="_blank" rel="noopener">get started&lt;/a> by &lt;a href="https://app.netdata.cloud/?utm_source=blog&amp;amp;utm_content=how_netdata_ml_works" target="_blank" rel="noopener">signing in to Netdata&lt;/a> and connecting a node. Once initial models have been trained (usually after the agent has about one hour of data, zero configuration needed), you&amp;rsquo;ll be able to start exploring in the &lt;a href="https://learn.netdata.cloud/docs/cloud/insights/anomaly-advisor" target="_blank" rel="noopener">Anomaly Advisor&lt;/a> tab of Netdata.&lt;/p></description></item><item><title>Anomaly rate in every chart</title><link>https://www.netdata.cloud/blog/anomaly-rate-in-every-chart/</link><pubDate>Thu, 23 Jun 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/anomaly-rate-in-every-chart/</guid><description>&lt;p>A month ago, we introduced unsupervised ML &amp;amp; Anomaly Detection in Netdata, the &lt;a href="https://www.netdata.cloud/blog/introducing-anomaly-advisor-unsupervised-anomaly-detection-in-netdata/">Anomaly Advisor&lt;/a>. Today, we’re happy to announce that we’re bringing anomaly rates to every chart in Netdata Cloud. Anomaly information is no longer limited to the Anomalies tab and will be accessible to you from the Overview and Single Node View tabs as well. This will make your troubleshooting journey easier, as you will have the anomaly rates for any metric available with a single click. Whichever metric or chart you&amp;rsquo;re exploring will be instant.&lt;/p></description></item><item><title>Metric Correlations on the Agent</title><link>https://www.netdata.cloud/blog/metric-correlations-on-the-agent/</link><pubDate>Wed, 15 Jun 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/metric-correlations-on-the-agent/</guid><description>&lt;p>As of &lt;a href="https://github.com/netdata/netdata/releases/tag/v1.35.0" target="_blank" rel="noopener">&lt;code>v1.35.0&lt;/code>&lt;/a> the Netdata Agent can now run &lt;a href="https://learn.netdata.cloud/docs/cloud/insights/metric-correlations" target="_blank" rel="noopener">Metric Correlations&lt;/a> (MC) itself. This means that, for nodes with MC enabled, the Metric Correlations feature just got a whole lot faster!&lt;/p>
&lt;!--truncate-->
&lt;p>The Netdata Metric Correlations feature uses a &lt;a href="https://en.wikipedia.org/wiki/Kolmogorov%E2%80%93Smirnov_test#Two-sample_Kolmogorov%E2%80%93Smirnov_test" target="_blank" rel="noopener">Two Sample Kolmogorov-Smirnov test&lt;/a> to look for which metrics have a significant distributional change around a highlighted window of interest. This can be useful when you are interested in short term &amp;ldquo;&lt;a href="https://en.wikipedia.org/wiki/Change_detection" target="_blank" rel="noopener">change detection&lt;/a>&amp;rdquo; and want to try answer the question &amp;ldquo;what else changed around this time?&amp;rdquo;.&lt;/p></description></item><item><title>Anomaly Advisor: Unsupervised Anomaly Detection</title><link>https://www.netdata.cloud/blog/introducing-anomaly-advisor-unsupervised-anomaly-detection-in-netdata/</link><pubDate>Thu, 26 May 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/introducing-anomaly-advisor-unsupervised-anomaly-detection-in-netdata/</guid><description>&lt;p>Today we are excited to launch one of our flagship ML assisted troubleshooting features in Netdata – the Anomaly Advisor.&lt;/p>
&lt;p>The Anomaly Advisor builds on earlier work to introduce unsupervised &lt;a href="https://github.com/netdata/netdata/blob/master/ml/README.md">anomaly detection&lt;/a> capabilities into the &lt;a href="https://www.netdata.cloud/agent/">Netdata Agent&lt;/a> from &lt;a href="https://github.com/netdata/netdata/releases/tag/v1.32.0">v1.32.0&lt;/a> onwards.&lt;/p>
&lt;h2 id="getting-started">Getting Started&lt;/h2>
&lt;p>Once you &lt;a href="https://learn.netdata.cloud/docs/configure/machine-learning#configuration">enable ML&lt;/a> on your nodes, each node will begin producing an &amp;ldquo;&lt;a href="https://learn.netdata.cloud/docs/configure/machine-learning#anomaly-bit---100--anomalous-0--normal">Anomaly Bit&lt;/a>&amp;rdquo; every second in addition to raw metric values. This anomaly bit will be 1 when the trained ML models consider recent raw data for a metric to look anomalous or 0 when things look &amp;rsquo;normal&amp;rsquo;. The Anomaly Advisor leverages this information to enable seamless space or room level anomaly detection out of the box with minimal configuration.&lt;/p></description></item><item><title>Monitoring without Cooperation: Kubernetes</title><link>https://www.netdata.cloud/blog/monitoring-without-cooperation-kubernetes/</link><pubDate>Fri, 20 May 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/monitoring-without-cooperation-kubernetes/</guid><description>&lt;!--truncate-->
&lt;iframe width="560" height="315" src="https://www.youtube.com/embed/J2kdSTRJzV4?si=Rqv4gpqz_wZKusaZ" title="YouTube video player" frameborder="0" allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" referrerpolicy="strict-origin-when-cross-origin" allowfullscreen>&lt;/iframe>
&lt;p>Imagine this: You are an engineer at a startup. You are responsible for keeping all the applications running smoothly and safely in production. At first, you have things under control, but soon enough things start getting more complex. The system has grown by hundreds of new server nodes, new services and containers are showing up seemingly every day, and you hear there’s a project to transition everything to Kubernetes (you’ve heard the term “cloud native” so often that the phrase shows up in your dreams).&lt;/p></description></item><item><title>Kubernetes Throttling Doesn’t Have To Suck. Let Us Help!</title><link>https://www.netdata.cloud/blog/kubernetes-throttling-doesnt-have-to-suck-let-us-help/</link><pubDate>Tue, 03 May 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/kubernetes-throttling-doesnt-have-to-suck-let-us-help/</guid><description>&lt;p>CPU limits are probably the most misunderstood concept in Kubernetes CPU resources allocation and management.&lt;/p>
&lt;!--truncate-->
&lt;p>A lot of engineers advise the use of CPU limits on every container as Kubernetes best practice. Unfortunately, as we will prove below, they are wrong: CPU limits should rarely be used, if used at all!&lt;/p>
&lt;p>But why? What are the reasons that even senior DevOps engineers with vast experience in the field advise the use of CPU limits?&lt;/p></description></item><item><title>Troubleshooting Alerts the Right Way: As a Team</title><link>https://www.netdata.cloud/blog/troubleshooting-alerts/</link><pubDate>Thu, 28 Apr 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/troubleshooting-alerts/</guid><description>&lt;!--truncate-->
&lt;p>At Netdata, we love two things more than anything else: &lt;/p>
&lt;ol>
 	&lt;li> Troubleshooting and,&lt;/li>
 	&lt;li> Making things easier for you! &lt;/li>
&lt;/ol>
Our goal is to make troubleshooting and monitoring as seamless as possible with the open-source Agent. This includes giving you pre-configured alerts so that you get notified immediately when a disruption occurs.
&lt;p>The Netdata Agent comes with over 250 pre-configured and optimized alerts. But we want you to be the master of your infrastructure monitoring by:&lt;/p></description></item><item><title>CNCF Live: Machine Learning Anomaly Detection</title><link>https://www.netdata.cloud/blog/cncf-live-power-up-your-machine-learning-automated-anomaly-detection/</link><pubDate>Wed, 27 Apr 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/cncf-live-power-up-your-machine-learning-automated-anomaly-detection/</guid><description>&lt;h2 id="join-us-live-to-talk-ml">Join Us Live to Talk ML&lt;/h2>
&lt;p>Join ML Lead Andrew Maguire and Product Manager Shyam Sreevalsan on the 23rd of June at 5pm UTC for the Netdata Machine Learning Meetup, which will be livestreamed on Youtube. In this session, we will demo and preview new and future features, along with a Q&amp;amp;A, and discuss the following topics:&lt;/p>
&lt;ul>
 	&lt;li>The role of Machine Learning in DevOps Infrastructure Monitoring &amp;amp; Troubleshooting&lt;/li>
 	&lt;li>The Netdata way of approaching Machine Learning&lt;/li>
 	&lt;li>Challenges of building Macvhine learning solutions that are useful and user friendly.&lt;/li>
&lt;/ul>
Feel free to ask questions or share ideas on &lt;a href="https://discord.gg/ZyeDHQTdaW">our Community Discord&lt;/a>, or &lt;a href="https://(https://www.meetup.com/netdata-infrastructure-monitoring-meetup-group/events/286243158/">RSVP to the event&lt;/a>. We look forward to seeing you then!
&lt;h2 id="cncf-live-power-up-your-machine-learning---automated-anomaly-detection">CNCF Live: Power up your machine learning - Automated anomaly detection&lt;/h2>
&lt;p>Our Analytics &amp;amp; ML lead Andrew Maguire recently had a chance to share our new &lt;a href="https://community.netdata.cloud/t/anomaly-advisor-beta-launch/2717">Anomaly Advisor&lt;/a> feature with the wider CNCF community. In his demonstration he did some light chaos engineering (using &lt;a href="https://www.gremlin.com/">Gremlin&lt;/a> and &lt;a href="https://wiki.ubuntu.com/Kernel/Reference/stress-ng">stress-ng&lt;/a>) to generate some real anomalies on his infrastructure and watch how it all played out in the Anomaly Advisor in Netdata Cloud.&lt;/p></description></item><item><title>The Netdata Way of Troubleshooting</title><link>https://www.netdata.cloud/blog/the-netdata-way-of-troubleshooting/</link><pubDate>Mon, 04 Apr 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/the-netdata-way-of-troubleshooting/</guid><description>&lt;p>Together with you, our fabulous community, Netdata is changing the way the world thinks of high fidelity monitoring - and we are gaining momentum.&lt;/p>
&lt;p>Our chief troublemaker and CEO, Costa Tsaousis,  is the pioneer and architect of this revolution that’s brewing in the monitoring and troubleshooting space.&lt;/p>
&lt;p>Watch him explain the &lt;strong>Netdata way of troubleshooting&lt;/strong>:&lt;/p>
&lt;iframe width="560" height="315" src="https://www.youtube.com/embed/ExjwwrgXvPg?si=IrsEp9PqbFFWiegM" title="YouTube video player" frameborder="0" allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" referrerpolicy="strict-origin-when-cross-origin" allowfullscreen>&lt;/iframe>
&lt;h2 id="why-did-the-world-need-a-new-infrastructure-monitoring-tool-0018">Why did the world need a new infrastructure monitoring tool? (00:18)&lt;/h2>
&lt;p>Every great hero needs an origin story. In this section, Costa explains the frustrating conditions in the infrastructure and monitoring space that lead to the conception, development, and subsequent success of the Netdata open-source Agent and Netdata Cloud.&lt;/p></description></item><item><title>Our Approach to Machine Learning</title><link>https://www.netdata.cloud/blog/our-approach-to-machine-learning/</link><pubDate>Fri, 25 Mar 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/our-approach-to-machine-learning/</guid><description>&lt;p>There is a lot of buzz in the world of machine learning (ML) and as a layperson it can be hard to keep up with it all. Therefore, we decided to write down some of our thoughts and musings on how &lt;b>we&lt;/b> are approaching ML at Netdata.&lt;/p>
&lt;!--truncate-->
&lt;h2 id="our-approach-to-machine-learning-ml">Our Approach to Machine Learning (ML)&lt;/h2>
&lt;p>We’ll touch on the current state of applied ML in industry in general, and zoom in on ML in the monitoring industry. We’ll discuss how we can leverage “good honest ML” to punch above our weight and add some useful and novel features for our users over the next few years.&lt;/p></description></item><item><title>Engineering Team Best Practices For Cloud Monitoring</title><link>https://www.netdata.cloud/blog/netdata-troubleshooting-show-engineering-team-best-practices-in-working-with-cloud/</link><pubDate>Wed, 02 Mar 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-troubleshooting-show-engineering-team-best-practices-in-working-with-cloud/</guid><description>&lt;!--truncate-->
&lt;figure class="wp-block-image size-large">&lt;img src="../wp-archive/uploads/2022/03/group_promo-1200x675.png" alt="" class="wp-image-16147"/>&lt;/figure>
&lt;h2 id="panel-engineering-team-best-practices-in-working-with-cloud">Panel: Engineering Team Best Practices in Working with Cloud&lt;/h2>
&lt;p>&lt;a href="https://youtube.com/playlist?list=PL-P-gAHfL2KMN-LHSmQFjpEwl31rQYvam" target="_blank">&lt;strong>The Netdata Troubleshooting Show&lt;/strong>&lt;/a>| Season 1: Episode 2 | Thursday, March 3rd at 9 am PST (UTC/GMT -8)&lt;/p>
&lt;p>&lt;strong>[&lt;a href="https://youtu.be/zY3DiRJ_DYc" target="_blank">Join Live&lt;/a>]&lt;/strong>For engineering teams, working in the cloud has never been easier, but also more complex. Enter a new world of remote working with distributed global teams, new tech challenges, new business realities, security, monitoring, and more.&lt;/p>
&lt;p>Join our dynamic panel of experts live&lt;strong>(&lt;/strong>&lt;a href="https://youtu.be/zY3DiRJ_DYc" target="_blank">&lt;strong>watch here&lt;/strong>&lt;/a>&lt;strong>)&lt;/strong>as they talk about engineering team best practices for those working in the cloud in 2022.&lt;/p></description></item><item><title>Netdata Meetup | Install &amp; Monitor From Scratch | Guide</title><link>https://www.netdata.cloud/blog/netdata-meetup-real-world-scenario-on-how-to-install-and-monitor-from-scratch/</link><pubDate>Wed, 02 Mar 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-meetup-real-world-scenario-on-how-to-install-and-monitor-from-scratch/</guid><description>&lt;!--truncate-->
&lt;p>&lt;img src="../wp-archive/uploads/2022/03/netdata-meetup-1-1200x676.png" alt="">&lt;/p>
&lt;p>Announcing Netdata Meetups: Virtual How-to Live Show – Join us Friday, March 4th&lt;/p>
&lt;p>&lt;a href="https://youtu.be/lBd0-TFJGAY">Join us live this Friday!&lt;/a> We are launching the brand new Netdata Meetups! Join us to learn more about Netdata. This how-to virtual series will help you go deeper with Netdata and as a bonus we will add in tips and tricks for monitoring and troubleshooting. We can’t wait to see you on Friday, March 4th at 9 am PST (GMT -8).&lt;/p></description></item><item><title>Netdata Cloud’s New Architecture</title><link>https://www.netdata.cloud/blog/netdata-clouds-new-architecture/</link><pubDate>Thu, 10 Feb 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-clouds-new-architecture/</guid><description>&lt;!--truncate-->
&lt;figure class="wp-block-image size-full">&lt;img class="wp-image-16154" src="../wp-archive/uploads/2022/03/New_arch_notification.jpeg" alt="" />&lt;/figure>
&lt;p>In version v1.32 of Netdata, we announced a remarkable new update that we are extremely proud of; Netdata Cloud now runs on the most reliable and stable backend that we’ve ever built. &lt;/p>
&lt;h2 id="migration-to-the-new-netdata-cloud-architecture">Migration to the new Netdata Cloud architecture&lt;/h2>
&lt;p>To give you the best experience of Netdata Cloud, we started migrating nodes running on the old architecture to the new one. Most users don’t have to take any action on their part. If you need to take action, you will see the pop-up window above in Netdata Cloud.  &lt;/p></description></item><item><title>Meet The Netdata Community: eBPF Hero Thiago</title><link>https://www.netdata.cloud/blog/meet-the-netdata-community-every-company-needs-an-ebpf-superhero-like-thiago/</link><pubDate>Fri, 04 Feb 2022 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/meet-the-netdata-community-every-company-needs-an-ebpf-superhero-like-thiago/</guid><description>&lt;!--truncate-->
&lt;p>&lt;img src="../wp-archive/uploads/2022/03/thiago-2.png" alt="">&lt;/p>
&lt;p>Our ongoing Meet the community series focuses on global Netdata community members. In this installment, learn about Netdata staff member and Software Engineer Thiago Marques, who is hard at work building an eBPF.plugin.&lt;/p>
&lt;h3 id="introduce-yourself-what-you-do-and-your-current-job-role-at-netdata">Introduce yourself, what you do, and your current job role at Netdata.&lt;/h3>
&lt;p>I am Thiago Marques, a C developer who works with the data collector team. My primary responsibility in the company is to develop eBPF.plugin.&lt;/p></description></item><item><title>Meet The Netdata Community: Rupok Chowdhury Protik</title><link>https://www.netdata.cloud/blog/meet-the-netdata-community-rupok-chowdhury-protik-takes-software-engineering-photography-to-new-heights/</link><pubDate>Tue, 07 Dec 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/meet-the-netdata-community-rupok-chowdhury-protik-takes-software-engineering-photography-to-new-heights/</guid><description>&lt;!--truncate-->
&lt;figure class="wp-block-image size-full">&lt;img src="../wp-archive/uploads/2022/03/Meet-the-community.png" alt="" class="wp-image-16164"/>&lt;/figure>
&lt;p>In our ongoing Meet the community series, we focus on global Netdata community members. In this installment, learn about Software Engineer and photographer Rupok Chowdhury Protik, including how they use Netdata.&lt;/p>
&lt;p>&lt;strong>Introduce yourself, what you do, and your current job role.&lt;/strong>&lt;/p>
&lt;p>Hi, I’m Rupok Chowdhury Protik. I’m a PHP Developer with a Master’s degree in Software Engineering. My stack is mainly LAMP/LEMP. I used to work as the Senior Web Application Developer in a multinational company and then later joined Incsub LLC, a WordPress-focused company with people from more than 50 countries. I am currently working in Incsub LLC as a DevOps Support Engineer.&lt;/p></description></item><item><title>Netdata Cloud Show: The Observability Panel</title><link>https://www.netdata.cloud/blog/netdata-cloud-show-the-observability-panel/</link><pubDate>Tue, 07 Dec 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-cloud-show-the-observability-panel/</guid><description>&lt;!--truncate-->
&lt;figure class="wp-block-image size-large">&lt;img class="wp-image-16169" src="../wp-archive/uploads/2022/03/Netdata-Cloud-Panel-1200x674.png" alt="" />&lt;/figure>
&lt;h2 id="in-2022-60-of-security-incidents-will-involve-third-partiesstrong-forresterstrong">“In 2022, 60% of security incidents will involve third parties.”&lt;strong>– Forrester&lt;/strong>&lt;/h2>
&lt;figure class="wp-block-embed is-type-rich is-provider-embed-handler wp-block-embed-embed-handler wp-embed-aspect-16-9 wp-has-aspect-ratio">
&lt;div class="wp-block-embed__wrapper">https://www.youtube.com/embed/kILpPCRlVD0&lt;/div>
&lt;/figure>
&lt;p>&lt;strong>This Thursday, December 9th at 1 PM EST (UTC-5)&lt;/strong>Join us live this Thursday for the first-ever Netdata Cloud Show, where our all-star panel will be discussing the latest trends around observability. &lt;/p>
&lt;p>This will be broadcast live on Netdata &lt;a href="https://youtu.be/kILpPCRlVD0">YouTube&lt;/a>, &lt;a href="https://twitter.com/linuxnetdata">Twitter&lt;/a>, &lt;a href="https://www.facebook.com/linuxnetdata/">Facebook&lt;/a>, and &lt;a href="https://discord.gg/kUk3nCmbtx">Discord&lt;/a>.&lt;/p>
&lt;p>&lt;strong>Observability Panel:&lt;/strong>&lt;/p>
&lt;ul>
&lt;li>&lt;strong>Boris Zaikin&lt;/strong>, Senior Software and Cloud Architect at Nordcloud. &lt;a href="https://twitter.com/boriszzn">@boriszzn&lt;/a>&lt;/li>
&lt;li>&lt;strong>Samir Behara&lt;/strong>, Platform Architect at EBSCO Industries, Inc. &lt;a href="https://twitter.com/samirbehara">@samirbehara&lt;/a>&lt;/li>
&lt;li>&lt;strong>Ralph Meijer&lt;/strong>, VP of Technology at Netdata. &lt;a href="https://twitter.com/ralphm">@ralphm&lt;/a>&lt;/li>
&lt;/ul>
&lt;p>&lt;strong>Topics Covered Include:&lt;/strong>&lt;/p></description></item><item><title>All-new Netdata Cloud Charts 2.0</title><link>https://www.netdata.cloud/blog/all-new-netdata-cloud-charts-2-0/</link><pubDate>Tue, 30 Nov 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/all-new-netdata-cloud-charts-2-0/</guid><description>&lt;!--truncate-->
&lt;figure class="wp-block-image size-large">&lt;img src="../wp-archive/uploads/2022/03/Netdata-Charts-2.0-1200x704.png" alt="" class="wp-image-16212"/>&lt;/figure>
&lt;p>Netdata excels in collecting, storing, and organizing metrics in out-of-the-box dashboards for powerful troubleshooting. We are now doubling down on this by transforming data into even more effective visualizations, helping you make the most sense out of all your metrics for increased observability.&lt;/p>
&lt;p>The new Netdata Charts provide a ton of useful information and we invite you to further explore our new charts from a design and development perspective. As always, it’s our goal to be as open and transparent as possible with our users on all things Netdata, including the ins and outs of how Netdata is built, why we make certain design decisions (driven by you of course!), and where we are heading as we continue to grow.&lt;/p></description></item><item><title>Meet The Netdata Community | Bastien, SysAdmin &amp; Dog Lover</title><link>https://www.netdata.cloud/blog/meet-the-netdata-community-learn-about-bastien-a-system-and-network-admin-who-loves-dogs/</link><pubDate>Tue, 30 Nov 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/meet-the-netdata-community-learn-about-bastien-a-system-and-network-admin-who-loves-dogs/</guid><description>&lt;!--truncate-->
&lt;figure class="wp-block-image size-full">&lt;img src="../wp-archive/uploads/2022/03/Screen-Shot-2021-11-30-at-3.10.06-PM.png" alt="" class="wp-image-16223"/>&lt;/figure>
&lt;p>&lt;strong>Tell people about yourself and what you do.&lt;/strong>&lt;/p>
&lt;p>My name is Bastien. I’m 28, a System and Network Admin, currently SRE in a team of two for my company. We’re working with AWS and Kubernetes to host our web apps.&lt;/p>
&lt;p>&lt;strong>How are you using Netdata and what do you like so far?&lt;/strong>&lt;/p>
&lt;p>I’m currently in the process of moving to Netdata to monitor our nine Kubernetes clusters and some standalone VMs. What I particularly love about Netdata is the simplicity of installation and configuration and the number of perfect default metrics and graphs.&lt;/p></description></item><item><title>How to extend the Geth collector</title><link>https://www.netdata.cloud/blog/how-to-extend-the-geth-collector/</link><pubDate>Mon, 13 Sep 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/how-to-extend-the-geth-collector/</guid><description>&lt;!--truncate-->
&lt;figure class="wp-block-image size-large">&lt;img src="../wp-archive/uploads/2022/03/Geth-collector-diagram-1-1200x796.png" alt="" class="wp-image-16282"/>&lt;/figure>
&lt;p>This is the the last of a 2-part blog post series regarding Netdata and Geth. If you missed the first, be sure to check it out &lt;a href="https://hackmd.io/J1x1WA-bR0a8gQeAmVdLFw" target="_blank" rel="noreferrer noopener">here&lt;/a>.&lt;/p>
&lt;p>Geth is short for Go-Ethereum and is the official implementation of the Ethereum Client in Go. Currently it’s one of the most widely used implementations and a core piece of infrastructure for the Ethereum ecosystem.&lt;/p>
&lt;p>With this proof of concept I wanted to showcase how easy it really is to gather data from any Prometheus endpoint and visualize them in Netdata. This has the added benefit of leveraging all the other features of Netdata, namely it’s per-second data collection, automatic deployment and configuration and superb system monitoring.&lt;/p></description></item><item><title>How to monitor the Geth node in under 5 minutes</title><link>https://www.netdata.cloud/blog/how-to-monitor-the-geth-node-in-under-5-minutes/</link><pubDate>Mon, 13 Sep 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/how-to-monitor-the-geth-node-in-under-5-minutes/</guid><description>&lt;!--truncate-->
&lt;figure class="wp-block-image size-large">&lt;img src="../wp-archive/uploads/2022/03/How-Geth-affect-the-CPU-charts-1200x629.png" alt="" class="wp-image-16229"/>&lt;/figure>
&lt;p>This piece is a blog post version of a workshop I gave at &lt;a href="https://ethcc.io/" target="_blank" rel="noreferrer noopener">EthCC&lt;/a> about monitoring an Ethereum Node using Netdata.&lt;/p>
&lt;p>&lt;strong>Disclaimer:&lt;/strong> Although we use Netdata, this guide is generic. We talk about metrics that can be surfaced by many other tools, such as Prometheus/Grafana or Datadog.&lt;/p>
&lt;p>The contents are as follows:&lt;/p>
&lt;ul>&lt;li class="">Introduction to Ethereum Nodes&lt;/li>&lt;li class="">What is Netdata&lt;/li>&lt;li class="">How to monitor a system that runs go-ethereum (Geth)&lt;/li>&lt;li class="">How to monitor go-ethereum (Geth)&lt;/li>&lt;/ul>
&lt;h2 id="ethereum-nodes">Ethereum Nodes&lt;/h2>
&lt;p>Running a node is no small feat, as it requires increasingly more and more resources to store the state of the blockchain and quickly process new transactions.&lt;/p></description></item><item><title>Root cause analysis using Metric Correlations</title><link>https://www.netdata.cloud/blog/root-cause-analysis-using-metric-correlations/</link><pubDate>Fri, 03 Sep 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/root-cause-analysis-using-metric-correlations/</guid><description>&lt;!--truncate-->
&lt;figure class="wp-block-image size-large">&lt;img src="../wp-archive/uploads/2022/03/Screen-Shot-2021-09-03-at-1.43.32-PM-1-1200x608.png" alt="" class="wp-image-16297"/>&lt;/figure>
&lt;p>As complexity of systems and applications continue to evolve and change, the number of metrics that need to be monitored grows in parallel. Whether you’re on a DevOps team, an SRE, or a developer building the code yourself, many of these components may be fragmented across your infrastructure, making it increasingly difficult to identify the root cause when experiencing downtime or abnormal behavior. To help solve this challenge, we built the &lt;a href="https://learn.netdata.cloud/docs/cloud/insights/metric-correlations">Metric Correlations&lt;/a> feature – an automated analysis tool that evaluates all your metrics to identify which have changed the most within a given period of interest.&lt;/p></description></item><item><title>How To Monitor Disks &amp; Filesystems With eBPF</title><link>https://www.netdata.cloud/blog/how-to-monitor-your-disks-and-filesystems-now-also-with-ebpf/</link><pubDate>Mon, 16 Aug 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/how-to-monitor-your-disks-and-filesystems-now-also-with-ebpf/</guid><description>&lt;!--truncate-->
&lt;div class="et_pb_module et_pb_text et_pb_text_0 et_pb_text_align_left et_pb_bg_layout_light">
&lt;div class="et_pb_text_inner">
&lt;img class="alignnone size-medium wp-image-16371" src="../wp-archive/uploads/2021/08/eBPF-monitoring-1536x932-1-600x364.png" alt="" width="600" height="364" />
&lt;h2 id="introduction-to-ebpf">Introduction to eBPF&lt;/h2>
&lt;/div>
&lt;/div>
&lt;div class="et_pb_module et_pb_text et_pb_text_1 et_pb_text_align_left et_pb_bg_layout_light">
&lt;div class="et_pb_text_inner">
&lt;p>Current IT monitoring software lacks the necessary metrics for minimizing downtime for systems and applications. Most provide system and application metrics but there is much more than this required for properly monitoring your infrastructure. With &lt;a title="eBPF" href="https://ebpf.io/" target="_blank" rel="noopener">eBPF&lt;/a> there is a technological advancement that allows monitoring software to provide rich information from the Linux kernel and present it. eBPF monitoring, specifically, provides a better understanding of what exactly is occurring on internal systems, which helps to identify where performance improvements can be made.&lt;/p></description></item><item><title>Netdata Is Launching Its Discord Server</title><link>https://www.netdata.cloud/blog/netdata-is-launching-its-discord-server/</link><pubDate>Tue, 22 Jun 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-is-launching-its-discord-server/</guid><description>&lt;!--truncate-->
&lt;p>It’s been a long time since our last community update, rest assured that we have been hard at work here at Netdata.&lt;/p>
&lt;p>Community building is hard, especially when you have such a venerable community like the one here at Netdata, where hundreds of contributors have contributed to creating one of the best monitoring solutions that exist.&lt;/p>
&lt;p>Last year we started to concentrate working on consolidating the community by integrating the various platforms where people come together to talk about Netdata. In that effort, we restarted our community forums in November, reworking the categories and creating a support channel, where community members help each other.&lt;/p></description></item><item><title>Netdata v1.31.0</title><link>https://www.netdata.cloud/blog/netdata-v1-31/</link><pubDate>Wed, 19 May 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-v1-31/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-medium wp-image-16402" src="../wp-archive/uploads/2021/05/v1.31.01-1-600x375.png" alt="" width="600" height="375" />
&lt;p>Give a warm welcome to Netdata v1.31.0, which features:&lt;/p>
&lt;ul>
 	&lt;li aria-level="1">&lt;strong>Re-packaged and redesigned dashboard&lt;/strong>: A more informational and feature-rich “frame” for your monitoring and troubleshooting sessions.&lt;/li>
 	&lt;li aria-level="1">&lt;strong>eBPF expands into the directory cache&lt;/strong>: Monitor whether your services or applications are properly using Linux’s memory management for the best performance and minimal disk I/O.&lt;/li>
 	&lt;li aria-level="1">&lt;strong>Machine learning-powered collectors&lt;/strong>: Detect anomalies using only your own data and minimal resource utilization on your monitored nodes.&lt;/li>
 	&lt;li aria-level="1">&lt;strong>An improved Netdata learning experience&lt;/strong>: A timeline of new content, refreshed visuals, and a newly-open sourced repository.&lt;/li>
&lt;/ul>
&lt;h2 id="h_1876248811621353710666">Re-packaged and redesigned dashboard&lt;/h2>
We re-packaged and redesigned portions of the dashboard to improve the overall experience. Part of this effort is better handling of dashboard code during installation—anyone using third-party packages (such as the Netdata Homebrew formula) will start seeing new features and the new designs starting today.
&lt;p>For those who aren’t using third-party packages (thank you!), your installation process will still get a little bit faster.&lt;/p></description></item><item><title>Kubernetes monitoring and troubleshooting made simple</title><link>https://www.netdata.cloud/blog/kubernetes-monitoring-troubleshooting/</link><pubDate>Wed, 05 May 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/kubernetes-monitoring-troubleshooting/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone wp-image-16419 size-large" src="../wp-archive/uploads/2022/03/kubernetes-monitoring-troubleshooting-1200x828.png" alt="" width="1200" height="828" />
&lt;p>Infrastructure monitoring was difficult enough when entire businesses ran off a few &lt;a href="https://www.netdata.cloud/academy/bare-metal-server/">bare metal servers&lt;/a> in a dusty, forgotten closet. Other IT infrastructure monitoring tools fell short, unable to provide complete and granular-enough metrics in real time, even when we were only dealing with a handful of systems responsible for running every part of the application stack. They were hard to configure, especially for the non-gurus out there, and didn’t provide the high-resolution metrics the gurus needed to make data-driven troubleshooting decisions.&lt;/p></description></item><item><title>Netdata 1.30.0</title><link>https://www.netdata.cloud/blog/netdata-1-30-0/</link><pubDate>Thu, 01 Apr 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-1-30-0/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-medium wp-image-16427" src="../wp-archive/uploads/2022/03/v1.30.0-600x338.png" alt="" width="600" height="338" />
&lt;p>We’re excited to introduce Netdata 1.30.0, which features:&lt;/p>
&lt;ul>
 	&lt;li>&lt;a href="https://staging-www.netdata.cloud/blog/release-1-30-0/#h_13754275791617303933429" target="_blank" rel="noopener">&lt;strong>ACLK-NG&lt;/strong>&lt;/a>: A new custom library for streaming metrics data on demand, written entirely in-house, that’s 4x faster than libmosquitto/libwebsockets.&lt;/li>
 	&lt;li>&lt;a href="https://staging-www.netdata.cloud/blog/release-1-30-0/#h_51169176821617303905806">&lt;strong>Opt-in product telemetry with PostHog&lt;/strong>&lt;/a>&lt;strong>:&lt;/strong> Goodbye, Google Analytics. Hello, self-hosted instance of PostHog.&lt;/li>
 	&lt;li>&lt;strong>&lt;a href="https://staging-www.netdata.cloud/blog/release-1-30-0/#h_612104678161617303939018">Deeper Linux kernel monitoring with eBPF&lt;/a>&lt;/strong>: Expanding our reach into the Linux kernel with page cache and synchronization syscall monitoring.&lt;/li>
 	&lt;li>&lt;strong>&lt;a href="https://staging-www.netdata.cloud/blog/release-1-30-0/#h_327502801221617303950492">Smarter preconfigured alarms&lt;/a>&lt;/strong>: Better (and less noisy) defaults, better information.&lt;/li>
 	&lt;li>&lt;strong>&lt;a href="https://staging-www.netdata.cloud/blog/release-1-30-0/#h_282794194271617303959644">Developer environment&lt;/a>&lt;/strong>: Contribute to Netdata via a Docker image and VSCode integration.&lt;/li>
 	&lt;li>&lt;strong>&lt;a href="https://staging-www.netdata.cloud/blog/release-1-30-0/#h_740997254311617303965795">Documentation improvements &amp;amp; tutorials&lt;/a>&lt;/strong>: Better standards for editing files and restarting Netdata, plus brand-new tutorials.&lt;/li>
&lt;/ul>
&lt;h2 id="h_13754275791617303933429">ACLK-NG&lt;/h2>
The ACLK-NG is a new, faster method of securely connecting a node running Netdata to Netdata Cloud. In our internal testing, it’s 4x faster than our previous implementation, which uses &lt;a href="https://github.com/netdata/mosquitto" target="_blank" rel="noopener">libmosquitto&lt;/a> and &lt;a href="https://github.com/warmcat/libwebsockets" target="_blank" rel="noopener">libwebsockets&lt;/a>.
&lt;table id="tablepress-10" class="tablepress tablepress-id-10 tablepress-responsive" >
&lt;thead>
&lt;tr class="row-1 odd">
&lt;th class="column-2" colspan="2">ACLK-NG&lt;/th>
&lt;th class="column-4" colspan="2">ACLK Legacy&lt;/th>
&lt;/tr>
&lt;/thead>
&lt;tbody class="row-hover">
&lt;tr class="row-2 even">
&lt;td class="column-1">&lt;/td>
&lt;td class="column-2">time (s)&lt;/td>
&lt;td class="column-3">MB/s&lt;/td>
&lt;td class="column-4">time (s)&lt;/td>
&lt;td class="column-5">MB/s&lt;/td>
&lt;/tr>
&lt;tr class="row-3 odd">
&lt;td class="column-1">Run 1&lt;/td>
&lt;td class="column-2">1.30&lt;/td>
&lt;td class="column-3">76.92&lt;/td>
&lt;td class="column-4">5.30&lt;/td>
&lt;td class="column-5">18.87&lt;/td>
&lt;/tr>
&lt;tr class="row-4 even">
&lt;td class="column-1">Run 2&lt;/td>
&lt;td class="column-2">1.31&lt;/td>
&lt;td class="column-3">76.34&lt;/td>
&lt;td class="column-4">5.29&lt;/td>
&lt;td class="column-5">18.90&lt;/td>
&lt;/tr>
&lt;tr class="row-5 odd">
&lt;td class="column-1">Run 3&lt;/td>
&lt;td class="column-2">1.38&lt;/td>
&lt;td class="column-3">72.46&lt;/td>
&lt;td class="column-4">5.27&lt;/td>
&lt;td class="column-5">18.98&lt;/td>
&lt;/tr>
&lt;tr class="row-6 even">
&lt;td class="column-1">Run 4&lt;/td>
&lt;td class="column-2">1.27&lt;/td>
&lt;td class="column-3">78.74&lt;/td>
&lt;td class="column-4">5.40&lt;/td>
&lt;td class="column-5">18.52&lt;/td>
&lt;/tr>
&lt;tr class="row-7 odd">
&lt;td class="column-1">Run 5&lt;/td>
&lt;td class="column-2">1.24&lt;/td>
&lt;td class="column-3">80.65&lt;/td>
&lt;td class="column-4">5.46&lt;/td>
&lt;td class="column-5">18.32&lt;/td>
&lt;/tr>
&lt;tr class="row-8 even">
&lt;td class="column-1">&lt;/td>
&lt;td class="column-2">&lt;strong>1.30&lt;/strong>&lt;/td>
&lt;td class="column-3">&lt;strong>77.02&lt;/strong>&lt;/td>
&lt;td class="column-4">&lt;strong>5.34&lt;/strong>&lt;/td>
&lt;td class="column-5">&lt;strong>18.72&lt;/strong>&lt;/td>
&lt;/tr>
&lt;/tbody>
&lt;/table>
With ACLK-NG enabled, you’ll get a snappier experience in Netdata Cloud, as there will be far less latency between requests for metrics and the subsequent response from individual nodes.
&lt;p>To enable ACLK-NG right now, update your nodes with the &lt;code>&amp;ndash;aclk-ng&lt;/code> option:&lt;/p></description></item><item><title>Container deployment showdown: Docker or Kubernetes?</title><link>https://www.netdata.cloud/blog/container-deployment-showdown-docker-or-kubernetes/</link><pubDate>Wed, 24 Mar 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/container-deployment-showdown-docker-or-kubernetes/</guid><description>&lt;!--truncate-->
&lt;div class="et_pb_module et_pb_text et_pb_text_0 et_pb_text_align_left et_pb_bg_layout_light">
&lt;div class="et_pb_text_inner">
&lt;img class="alignnone wp-image-16433 size-large" src="../wp-archive/uploads/2022/03/Kubernetes_vs_Docker-1200x828.png" alt="" width="1200" height="828" />
&lt;p>Monitoring the current state and performance of applications is critical for IT Ops and DevOps teams alike. Understanding the health of an application is one of the most effective ways of anticipating potential bottlenecks or slowdowns, yet it’s one of the largest challenges faced by many organizations that build and deploy software. This is largely due to applications’ distributed and diversified nature. A single outage has the potential to interrupt entire processes that, at times, can interfere with business as a whole and result in a negative effect on the bottom line.&lt;/p></description></item><item><title>5 DevOps best practices to reinforce with monitoring tools</title><link>https://www.netdata.cloud/blog/devops-best-practices-monitoring-tools/</link><pubDate>Thu, 18 Feb 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/devops-best-practices-monitoring-tools/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-medium wp-image-16441" src="../wp-archive/uploads/2022/03/devops-best-practices-monitoring-tools-v2-600x414.png" alt="" width="600" height="414" />
&lt;p>As part of a modern software development team, you’re asked to do a lot. You’re supposed to build faster, release more frequently, crush bugs, and integrate testing suites along the way. You’re supposed to implement and practice a strong DevOps culture, &lt;a href="https://github.com/upgundecha/howtheysre" target="_blank" rel="noopener noreferrer">read entire novels&lt;/a> about SRE best practices, go &lt;a href="https://en.wikipedia.org/wiki/Agile_software_development" target="_blank" rel="noopener noreferrer">agile&lt;/a>, or add a bunch of Scrum ceremonies to everyone’s calendar. Every week, the industry recommends that you “&lt;a href="https://devops.com/devops-shift-left-avoid-failure/" target="_blank" rel="noopener noreferrer">shift-left&lt;/a>” another part of the &lt;a href="https://staging-www.netdata.cloud/blog/agile-static-analysis/" target="_blank" rel="noopener noreferrer">DevOps pipeline&lt;/a>, to the point where you’re supposed to handle everything from unit testing to production deployment optimization from day one.&lt;/p></description></item><item><title>Introduction to StatsD</title><link>https://www.netdata.cloud/blog/introduction-to-statsd/</link><pubDate>Wed, 03 Feb 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/introduction-to-statsd/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-medium wp-image-16449" src="../wp-archive/uploads/2022/03/StatsD-600x414.png" alt="" width="600" height="414" />
&lt;p>StatsD is an industry-standard technology stack for monitoring applications and instrumenting any piece of software to deliver custom metrics. The StatsD architecture is based on delivering the metrics via UDP packets from any application to a central statsD server. Although the original StatsD server was written in Node.js, there are many implementations today, with Netdata being one of them.&lt;/p>
&lt;p>StatsD makes it easier for you to instrument your applications, delivering value around three main pillars: open-source, control, and modularity. That’s a real windfall for full-stack developers who need to code quickly, troubleshoot application issues on the fly, and often don’t have the necessary background knowledge to use complex monitoring platforms.&lt;/p></description></item><item><title>Actionable Alerts With Fewer False Positives</title><link>https://www.netdata.cloud/blog/actionable-intelligent-alerts/</link><pubDate>Thu, 21 Jan 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/actionable-intelligent-alerts/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-medium wp-image-16460" src="../wp-archive/uploads/2022/03/intelligent-alarms-600x413.png" alt="" width="600" height="413" />
&lt;p>Think about any sport or competitive activity, whether that’s football or a spelling bee. They always feature at least one person who acts as a moderator, referee, or judge. With their domain expertise, this person watches everyone’s behavior and constantly compares that against a set of rules. If someone crosses that threshold, they blow a whistle or throw up a flag. They are, in effect, saying that things have gone from &lt;strong>OK&lt;/strong> to &lt;strong>not OK&lt;/strong>.&lt;/p></description></item><item><title>Four key metrics for responding to IT incidents and failures</title><link>https://www.netdata.cloud/blog/four-key-metrics-for-responding-to-it-incidents-and-failures/</link><pubDate>Thu, 07 Jan 2021 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/four-key-metrics-for-responding-to-it-incidents-and-failures/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone wp-image-16474 size-large" src="../wp-archive/uploads/2021/01/DevOps-Metrics-1200x862.png" alt="" width="1200" height="862" />
&lt;div class="et_pb_module et_pb_text et_pb_text_0 et_pb_text_align_left et_pb_bg_layout_light">
&lt;div class="et_pb_text_inner">
&lt;p>If you’re a veteran in this space, you probably understand the many incident response metrics and concepts, along with the many (at times exasperating) acronyms. For those new to the space, or even those with years of experience, the &lt;a title="terminology" href="https://en.wikipedia.org/wiki/List_of_computing_and_IT_abbreviations" target="_blank" rel="noopener noreferrer">terminology&lt;/a> is often overwhelming.&lt;/p>
&lt;p>If you’re one of those people who’s struggling to navigate through the world of DevOps metrics, we’ve created this article for you. In this post, we’ll cover four main incident response metrics: MTTA, MTTR, MTBF, and MTTF. Learn what these acronyms mean, how you can use them, and how you can tie these KPIs into your Netdata experience.&lt;/p></description></item><item><title>Netdata Year in Review 2020</title><link>https://www.netdata.cloud/blog/netdata-year-in-review-2020/</link><pubDate>Fri, 18 Dec 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-year-in-review-2020/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-medium wp-image-16482" src="../wp-archive/uploads/2022/03/Roadmap-Header-600x322.png" alt="" width="600" height="322" />
&lt;p>Looking back at the unprecedented challenges we faced together in 2020, we’d like to extend our thanks to the community of people who have continued to work towards Netdata’s mission of simplifying monitoring and troubleshooting for everyone. Let’s review some of this year’s highlights.&lt;/p>
&lt;table>
&lt;tbody>
&lt;tr>
&lt;td width="50%">
&lt;table>
&lt;tbody>
&lt;tr>
&lt;td>
&lt;h3>Community&lt;/h3>
&lt;/td>
&lt;td>More than &lt;a title="https://github.com/netdata/netdata/graphs/contributors" href="https://github.com/netdata/netdata/graphs/contributors" target="_blank" rel="noopener noreferrer">400 contributors&lt;/a> have helped us grow.&lt;/td>
&lt;/tr>
&lt;tr>
&lt;td>Nearly 50,000 GitHub &lt;a title="stargazers" href="https://github.com/netdata/netdata/stargazers" target="_blank" rel="noopener noreferrer">stargazers&lt;/a> follow our progress.&lt;/td>
&lt;td>Our community on &lt;a title="GitHub" href="https://github.com/netdata/netdata/" target="_blank" rel="noopener noreferrer">GitHub&lt;/a> and our &lt;a title="forums" href="https://community.netdata.cloud/" target="_blank" rel="noopener noreferrer">forums&lt;/a> has grown to 4,000 strong.&lt;/td>
&lt;/tr>
&lt;/tbody>
&lt;/table>
&lt;/td>
&lt;td width="50%">&lt;img class="wp-image-16480 aligncenter" src="../wp-archive/uploads/2022/03/community.png" alt="" width="227" height="225" />&lt;/td>
&lt;/tr>
&lt;/tbody>
&lt;/table>
&amp;nbsp;
&lt;table>
&lt;tbody>
&lt;tr>
&lt;td width="50%">&lt;img class="size-full wp-image-16492 aligncenter" src="../wp-archive/uploads/2020/12/stargazers.png" alt="" width="225" height="226" />&lt;/td>
&lt;td width="50%">
&lt;table>
&lt;tbody>
&lt;tr>
&lt;td>
&lt;h3>Netdata Cloud&lt;/h3>
&lt;/td>
&lt;td>Launched in May, now with more than 25,000 users &lt;a title="registered" href="https://app.netdata.cloud/" target="_blank" rel="noopener noreferrer">registered&lt;/a>.&lt;/td>
&lt;/tr>
&lt;tr>
&lt;td>Continuous delivery of new features included custom dashboards, Metric Correlations, overview page, and centralized alarm notifications.&lt;/td>
&lt;/tr>
&lt;/tbody>
&lt;/table>
&lt;/td>
&lt;/tr>
&lt;/tbody>
&lt;/table>
&lt;table>
&lt;tbody>
&lt;tr>
&lt;td width="50%">
&lt;table>
&lt;tbody>
&lt;tr>
&lt;td>
&lt;h3>Netdata Agent&lt;/h3>
&lt;/td>
&lt;td>9 dot &lt;a title="releases" href="https://github.com/netdata/netdata/releases" target="_blank" rel="noopener noreferrer">releases&lt;/a>, with support for eBPF, Prometheus metrics, &amp;amp; k8s service discovery.&lt;/td>
&lt;/tr>
&lt;tr>
&lt;td>More than 300 bug fixes, nearly 1,500 closed &lt;a title="issues" href="https://github.com/netdata/netdata/issues" target="_blank" rel="noopener noreferrer">issues&lt;/a>, and almost 3,200 commits.&lt;/td>
&lt;td>Nearly 400 new features or &lt;a title="improvements" href="https://github.com/netdata/netdata/" target="_blank" rel="noopener noreferrer">improvements&lt;/a>.&lt;/td>
&lt;/tr>
&lt;/tbody>
&lt;/table>
&lt;/td>
&lt;td width="50%">&lt;img class="wp-image-16499 size-full aligncenter" src="../wp-archive/uploads/2020/12/Github-1.png" alt="" width="256" height="256" />&lt;/td>
&lt;/tr>
&lt;/tbody>
&lt;/table>
&amp;nbsp;
&lt;table>
&lt;tbody>
&lt;tr>
&lt;td width="50%">&lt;img class="wp-image-16501 size-full aligncenter" src="../wp-archive/uploads/2020/12/funding-2.png" alt="" width="225" height="225" />&lt;/td>
&lt;td width="50%">
&lt;table>
&lt;tbody>
&lt;tr>
&lt;td>
&lt;h3>Company&lt;/h3>
&lt;/td>
&lt;td>&lt;a title="Raised" href="https://staging-www.netdata.cloud/news/netdata-extends-series-a-funding/">Raised&lt;/a> $14.2M, extending the total raised to $31M.&lt;/td>
&lt;/tr>
&lt;tr>
&lt;td>Recognized in &lt;a title="Forbes Cloud 100 Rising Stars 2020" href="https://staging-www.netdata.cloud/blog/forbes-cloud-100-rising-stars-2020/">Forbes Cloud 100 Rising Stars 2020&lt;/a>, &lt;a title="2020 Stratus Awards" href="https://staging-www.netdata.cloud/news/">2020 Stratus Awards&lt;/a> as Cloud Disruptor.&lt;/td>
&lt;/tr>
&lt;/tbody>
&lt;/table>
&lt;/td>
&lt;/tr>
&lt;/tbody>
&lt;/table>
&amp;nbsp;
&lt;p> &lt;/p></description></item><item><title>Centralize Infrastructure With Alarm Notifications</title><link>https://www.netdata.cloud/blog/cloud-alarm-notifications/</link><pubDate>Thu, 17 Dec 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/cloud-alarm-notifications/</guid><description>&lt;!--truncate-->
&lt;div class="et_pb_module et_pb_text et_pb_text_0 et_pb_text_align_left et_pb_bg_layout_light">
&lt;div class="et_pb_text_inner">
&lt;img class="alignnone wp-image-16509 size-large" src="../wp-archive/uploads/2020/12/Central-Alarm-Notifications-1200x828.png" alt="" width="1200" height="828" />
&lt;p>Netdata is architected on every level, across both the open-source Netdata Agent and Netdata Cloud, to help you own every layer of your monitoring experience. With this design, all metrics data collected by the Netdata Agent stays distributed on your node, but you also leverage Netdata Cloud’s dashboards and multi-node visualizations to view the health and performance of an entire infrastructure from a single application.&lt;/p></description></item><item><title>Community Update: Discourse, Community Efforts</title><link>https://www.netdata.cloud/blog/community-update-discourse/</link><pubDate>Wed, 02 Dec 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/community-update-discourse/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone wp-image-16522 size-full" src="../wp-archive/uploads/2022/03/Community-update_-Discourse-community-efforts.png" alt="" width="681" height="470" />
&lt;p>Open source and community have always been in the DNA of Netdata, with the Agent starting as a very popular open-source project. Since then, a lot has changed, with Netdata maturing into a company, and the Netdata Agent finding its place as an open-source project in a wider offering that redesigns the monitoring experience from the ground up.&lt;/p>
&lt;p>While we had a very active &lt;a href="https://github.com/netdata/netdata/" target="_blank" rel="noopener noreferrer">GitHub repository&lt;/a>, with the majority of the Netdata Agent’s original team actively moderating the discussions and talking with users, we started a more concentrated initiative to manage our community in spring 2020. We launched our first forum using &lt;a href="https://nodebb.org/" target="_blank" rel="noopener noreferrer">NodeBB&lt;/a>, a great open-source project, and the community grew substantially, outgrowing the forum software. At the same time, we were able to identify areas of friction in the community journey. Removing that friction became a centerpiece of our strategy during Q3 2020.&lt;/p></description></item><item><title>StackPulse: Automated Incident Remediation Workflow</title><link>https://www.netdata.cloud/blog/netdata-stackpulse-remediation/</link><pubDate>Wed, 02 Dec 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-stackpulse-remediation/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16535" src="../wp-archive/uploads/2022/03/netdata-stackpulse-1200x826.png" alt="" width="1200" height="826" />
&lt;p>Teams of all types use Netdata to monitor the health of their nodes with preconfigured alarms and real-time interactive visualizations, and when incidents happen, they troubleshoot issues with thousands of per-second metrics on &lt;a href="https://staging-www.netdata.cloud/cloud/" target="_blank" rel="noopener noreferrer">Netdata Cloud&lt;/a>. But based on the complexity of the team and the infrastructure they monitor, some parts of their &lt;a href="https://staging-www.netdata.cloud/incident-management/" target="_blank" rel="noopener noreferrer">incident management&lt;/a>, such as pre-planned communication and escalation processes, or even automated remediation, need to happen outside of the Netdata ecosystem.&lt;/p></description></item><item><title>What is DevOps?</title><link>https://www.netdata.cloud/blog/what-is-devops/</link><pubDate>Thu, 19 Nov 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/what-is-devops/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-full wp-image-16550" src="../wp-archive/uploads/2022/03/what-is-devops.png" alt="" width="1024" height="600" />
&lt;p>In software development, it’s important to have a team dedicated to ensuring all systems and applications maintain maximum performance and uptime. Establishing processes that limit system and application slowdowns and outages while expediting the product release process is often done through a developer operations team, also known as &lt;em>dev ops&lt;/em> or &lt;em>DevOps&lt;/em>.&lt;/p>
&lt;p>DevOps teams are responsible for improving communication and collaboration across engineering teams to increase an organization’s ability to efficiently deliver products and services that serve the end-user. In this post, we’ll describe the different elements of DevOps, including what DevOps is, how it works, and how Netdata helps DevOps teams succeed.&lt;/p></description></item><item><title>Community Repository: Consul, Ansible &amp; ML Recipes</title><link>https://www.netdata.cloud/blog/welcome-to-netdatas-community-repository-consul-ansible-ml/</link><pubDate>Wed, 11 Nov 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/welcome-to-netdatas-community-repository-consul-ansible-ml/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16555" src="../wp-archive/uploads/2022/03/netdata-community-repository-1200x825.png" alt="" width="1200" height="825" />
&lt;p>On our journey to democratize monitoring, we are proud to have open source at the core of both our products and our company values. What started as a project out of frustration for lack of existing alternatives (see &lt;a href="https://www.rexfeng.com/blog/2016/01/anger-driven-development/" target="_blank" rel="noopener noreferrer">anger-driven development&lt;/a>), quickly became one of the most starred open-source projects on all of GitHub.&lt;/p>
&lt;p>Fast-forward a couple of years later, and the Netdata Agent, our open-source monitoring agent, is maturing as the best single-node monitoring experience, offering unparalleled efficiency and thousands of metrics, per-second. At the same time, we have gathered a considerable community on our GitHub repository and new forums.&lt;/p></description></item><item><title>How Netdata gets you from 0 to monitoring in minutes</title><link>https://www.netdata.cloud/blog/how-netdata-gets-you-from-0-to-monitoring-in-minutes/</link><pubDate>Wed, 11 Nov 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/how-netdata-gets-you-from-0-to-monitoring-in-minutes/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16562" src="../wp-archive/uploads/2022/03/0conf-1200x826.png" alt="" width="1200" height="826" />
&lt;p>Netdata is zero-configuration monitoring. It’s a principle that we’ve stood behind since the project’s beginning, when it was only our CEO Costa trying to solve a &lt;a href="https://staging-www.netdata.cloud/blog/why-netdata-is-free/">“painful, real-world problem,”&lt;/a> and it’s one we stand by today. Our insistence on zero-configuration guides every product decision we make, every grooming process, and every React component our frontend teams design.&lt;/p>
&lt;p>In fact, zero-configuration is the exact reason why Netdata’s dashboard is &lt;a href="https://staging-www.netdata.cloud/blog/netdata-agent-dashboard/">open and accessible by default&lt;/a>. It’s how Netdata gets you from 0 monitoring to thousands of metrics, collected every second and visualized in real time, in a matter of minutes.&lt;/p></description></item><item><title>Netdata’s dashboard: open by default and secure by design</title><link>https://www.netdata.cloud/blog/netdata-agent-dashboard/</link><pubDate>Wed, 28 Oct 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-agent-dashboard/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16569" src="../wp-archive/uploads/2022/03/netdata-dashboard-open-secure-min-1200x826.png" alt="" width="1200" height="826" />
&lt;p>Let’s talk through a scenario: You have a Linux-based VM running on DigitalOcean (aka a Droplet), and you install Netdata on it using our &lt;a href="https://learn.netdata.cloud/docs/get#install-the-netdata-agent" target="_blank" rel="noopener noreferrer">recommended kickstart script&lt;/a>. As the installation process winds down, the Droplet starts up the Netdata Agent’s web server and serves the local Agent web dashboard on port 19999. You navigate to the dashboard using your browser of choice, check out per-second metrics updating in real time in a few of the hundreds of preconfigured visualizations, then realize…&lt;/p></description></item><item><title>Real-Time Infrastructure Monitoring Now In Netdata Cloud</title><link>https://www.netdata.cloud/blog/bringing-rich-and-real-time-infrastructure-monitoring-to-netdata-cloud/</link><pubDate>Thu, 22 Oct 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/bringing-rich-and-real-time-infrastructure-monitoring-to-netdata-cloud/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16578" src="../wp-archive/uploads/2022/03/Correlation_charts-1200x830.png" alt="" width="1200" height="830" />
&lt;p>The Netdata Agent is well-equipped to solve monitoring and troubleshooting challenges for single nodes. We love that the Agent is so valuable to our users, but Netdata Cloud is designed for infrastructure monitoring. That’s why we’re working so hard to offer even more capabilities and help users monitor and troubleshoot infrastructures of all sizes, entirely for free!&lt;/p>
&lt;p>With the new Cloud Overview, you get every real-time chart and metric you need to understand the status of your infrastructure, explore, and troubleshoot, in a single view. We designed the Overview on one existing and beloved feature and another entirely new one that we’re very excited to launch for the first time.&lt;/p></description></item><item><title>The reality of Netdata’s long-term metrics storage database</title><link>https://www.netdata.cloud/blog/the-reality-of-netdatas-long-term-metrics-storage-database/</link><pubDate>Mon, 12 Oct 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/the-reality-of-netdatas-long-term-metrics-storage-database/</guid><description>&lt;!--truncate-->
&lt;p>The perception that Netdata is only capable of short-term metrics storage is a myth. It’s a pervasive myth we still see in blog posts and through community engagement, despite it being false for more than a year.&lt;/p>
&lt;p> &lt;/p>
&lt;p>However, like all myths, this one on metrics storage began with a kernel of truth. When Netdata first flourished as an &lt;a title="open-source project" href="https://github.com/netdata/netdata" target="_blank" rel="noopener noreferrer">open-source project&lt;/a> in 2017 and 2018, the default metrics database was RAM-only. You could configure this database’s size, but for many users, that size was limited by the amount of RAM they were willing to allocate for metrics storage. We also kept the default value low to ensure Netdata worked efficiently on all hardware and a variety of operating systems.&lt;/p></description></item><item><title>Software Extensibility Is Key To Adoption</title><link>https://www.netdata.cloud/blog/software-extensibility-is-key-to-adoption/</link><pubDate>Fri, 25 Sep 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/software-extensibility-is-key-to-adoption/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-full wp-image-16613" src="../wp-archive/uploads/2022/03/Software-Extensibility-Blog-Post.png" alt="" width="969" height="638" />
&lt;p>As with most commercial products, software today is mass produced for reasons of simple economics. But making the same software work for as many people as possible while also meeting the unique requirements different people and organizations have is a challenging task.&lt;/p>
&lt;p>The strategies used today to provide extensibility at cost and at the level required by various enterprises are very similar to the ones used in the 80s and 90s, when PCs first took off and captured an enormous amount of the home computing market, and when open source software started gaining popularity.&lt;/p></description></item><item><title>Investing in Netdata: a growth story</title><link>https://www.netdata.cloud/blog/investing-in-netdata-a-growth-story/</link><pubDate>Tue, 22 Sep 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/investing-in-netdata-a-growth-story/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16618" src="../wp-archive/uploads/2022/03/Series-A-funding-extension-1200x571.png" alt="" width="1200" height="571" />
&lt;p>I’m excited to announce an &lt;a title="extension to Netdata’s series A funding" href="https://staging-www.netdata.cloud/news/netdata-extends-series-a-funding/" target="_blank" rel="noopener noreferrer">extension to Netdata’s series A funding &lt;/a>in the amount of $14.2M, bringing the total amount of funding to $31M. We’re thrilled to share the news; the additional funding will help us continue building the future of health monitoring and performance troubleshooting. In case you missed it, our mission is to &lt;a title="redefine infrastructure monitoring" href="https://staging-www.netdata.cloud/blog/redefining-monitoring-netdata/" target="_blank" rel="noopener noreferrer">redefine infrastructure monitoring&lt;/a>. Our unique approach to building the right solution with and for the community is no easy task.&lt;/p></description></item><item><title>Metric Correlations: Detect Patterns &amp; Anomalies</title><link>https://www.netdata.cloud/blog/netdata-cloud-metric-correlations/</link><pubDate>Wed, 16 Sep 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-cloud-metric-correlations/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16623" src="../wp-archive/uploads/2022/03/Cloud-Correlations@2x-1200x826.png" alt="" width="1200" height="826" />
&lt;p>Today, we are excited to launch our first Netdata Cloud Insights feature, Metric Correlations, developed for discovering underlying issues more quickly and identifying the root cause more efficiently. Read on to learn more about our approach to developing this new feature, how it works, and the many benefits you’ll find incorporating this into your team’s troubleshooting workflow.&lt;/p>
&lt;h2>Some background&lt;/h2>
Let’s start with a bit of a disclaimer. It seems machine learning (ML) (or “Artificial Intelligence,” if you are looking for more LinkedIn likes) has gone mainstream in the last few years, and we are probably by now somewhere near the “Peak of Inflated Expectations” on the &lt;a title="hype cycle" href="https://en.wikipedia.org/wiki/Hype_cycle" target="_blank" rel="noopener noreferrer">hype cycle&lt;/a>. It is in this context that we want to be clear about what our goals are in this space and our approach to releasing data-driven features that draw on techniques from statistics and ML. In short, we want to be clear, open, realistic, and avoid buzzwords at all costs!
&lt;p>Over the next 12 months, we are hoping to begin building a layer of intelligence&lt;sup>&lt;a href="https://staging-www.netdata.cloud/blog/netdata-cloud-metric-correlations/#1">1&lt;/a>&lt;/sup> throughout Netdata (both Cloud and Agent) to assist with “&lt;a title="human in the loop" href="https://hai.stanford.edu/blog/humans-loop-design-interactive-ai-systems" target="_blank" rel="noopener noreferrer">human in the loop&lt;/a>” troubleshooting, mainly to help users more easily surface slowdowns, anomalies, or other issues and lower your &lt;a title="cognitive load" href="https://en.wikipedia.org/wiki/Cognitive_load" target="_blank" rel="noopener noreferrer">cognitive load&lt;/a>&lt;sup>&lt;a href="https://staging-www.netdata.cloud/blog/netdata-cloud-metric-correlations/#2">2&lt;/a>&lt;/sup> as you troubleshoot using Netdata. Simply put, we’re working to streamline your mean time to resolution (MTTR).&lt;/p></description></item><item><title>Netdata Agent v1.25 &amp; Cloud Enhancements</title><link>https://www.netdata.cloud/blog/release-1-25/</link><pubDate>Wed, 16 Sep 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/release-1-25/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16635" src="../wp-archive/uploads/2022/03/Agent-Release-v1.25@2x-1200x600.png" alt="" width="1200" height="600" />
&lt;p>The &lt;a title="v1.25.0" href="https://github.com/netdata/netdata/releases" target="_blank" rel="noopener noreferrer">v1.25.0&lt;/a> release of the Netdata Agent delivers on our commitment to make our metrics collection, visualization, and troubleshooting platform more stable and usable. We enhanced our recently-added &lt;a title="Prometheus collector" href="https://learn.netdata.cloud/docs/agent/collectors/go.d.plugin/modules/prometheus" target="_blank" rel="noopener noreferrer">Prometheus collector&lt;/a> with user-configurable filtering and grouping, made dramatic improvements to the reliability of the Agent-Cloud link that streams metrics on-demand to your browser when you use &lt;a title="Netdata Cloud" href="https://app.netdata.cloud/" target="_blank" rel="noopener noreferrer">Netdata Cloud&lt;/a>, and more.&lt;/p>
&lt;p>Let’s jump in and look at each improvement.&lt;/p></description></item><item><title>Netdata named to the Forbes Cloud 100 Rising Stars</title><link>https://www.netdata.cloud/blog/forbes-cloud-100-rising-stars-2020/</link><pubDate>Wed, 16 Sep 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/forbes-cloud-100-rising-stars-2020/</guid><description>&lt;!--truncate-->
&lt;img class=" wp-image-16648 alignleft" src="../wp-archive/uploads/2022/03/Cloud1002020-RisingStars-SMALL.png" alt="" width="390" height="471" />
We’re excited to announce that we’ve been named to the&lt;strong> &lt;a title="Forbes 2020 Cloud 100 Rising Stars" href="https://www.forbes.com/sites/kenrickcai/2020/09/16/cloud-100-rising-stars-2020/" target="_blank" rel="noopener noreferrer">Forbes 2020 Cloud 100 Rising Stars&lt;/a>.&lt;/strong> This is a list of the top 100 private cloud companies in the world, published by Forbes in partnership with &lt;a title="Bessemer Venture Partners" href="https://www.bvp.com/" target="_blank" rel="noopener noreferrer">Bessemer Venture Partners&lt;/a> and &lt;a title="Salesforce Ventures" href="https://www.salesforce.com/company/ventures/" target="_blank" rel="noopener noreferrer">Salesforce Ventures&lt;/a>. The 20 Rising Stars represent young, high-growth and category-leading cloud companies who are poised to join the &lt;a title="Cloud 100" href="https://www.forbes.com/cloud100/" target="_blank" rel="noopener noreferrer">Cloud 100&lt;/a> ranks.
We are extremely honored to be recognized amongst our most-promising peers. This is a testament to our momentum we’ve built in partnership with our passionate community of users and contributors worldwide, as well as a validation of our community-first, open-source approach to democratizing infrastructure monitoring and troubleshooting.
The major milestones we’ve hit along the way this year include more than 5,000 new users added each day, more than 3 million users total worldwide, and nearly 50,000 GitHub stars, making Netdata the fourth most starred project in the Cloud Native Computing Foundation landscape. We’ve also launched our monitoring service, Netdata Cloud, which has seen very strong growth since its introduction in May, with more than 15,000 users registered so far. We couldn’t have done it without you!
The Forbes 2020 &lt;a title="Cloud 100" href="https://www.forbes.com/cloud100" target="_blank" rel="noopener noreferrer">Cloud 100&lt;/a> and 20 Rising Stars lists will also appear in the September 2020 issue of &lt;em>Forbes&lt;/em> magazine.</description></item><item><title>The Netdata Community Powered by NodeBB</title><link>https://www.netdata.cloud/blog/the-netdata-community-powered-by-nodebb/</link><pubDate>Thu, 13 Aug 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/the-netdata-community-powered-by-nodebb/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16658" src="../wp-archive/uploads/2022/03/NodeBB-1-1200x877.png" alt="" width="1200" height="877" />
&lt;p>We recently adopted &lt;a title="NodeBB" href="https://nodebb.org/" target="_blank" rel="noopener noreferrer">NodeBB&lt;/a> as our software of choice for building &lt;a title="the Netdata Community" href="https://community.netdata.cloud/" target="_blank" rel="noopener noreferrer">the Netdata Community&lt;/a>. We have &lt;a title="many good reasons" href="https://staging-www.netdata.cloud/blog/the-netdata-community/" target="_blank" rel="noopener noreferrer">many good reasons&lt;/a> for why we wanted to provide our community with a proper home online, but I wanted to cover some of the technical reasons for choosing NodeBB for our platform, and the many parallels between the NodeBB and Netdata projects, which was certainly a driving force behind this decision.&lt;/p></description></item><item><title>Release 1.24: Prometheus Collector &amp; Multi-Host DB</title><link>https://www.netdata.cloud/blog/release-1-24/</link><pubDate>Mon, 10 Aug 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/release-1-24/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-full wp-image-16663" src="../wp-archive/uploads/2022/03/1.24-release-2.png" alt="" width="683" height="469" />
&lt;p>The v1.24.0 release of the Netdata Agent brings enhancements to the breadth of metrics we collect with a new Prometheus/OpenMetrics collector and enhanced storage and querying with a new multi-host database mode. Let’s take a look at each of these enhancements.&lt;/p>
&lt;h2>Instantly access thousands of metrics with the new Prometheus/OpenMetrics collector&lt;/h2>
This release broadens our commitment to open standards, interoperability, and extensibility with a new generic Prometheus collector that works seamlessly with any application that makes its metrics available in the &lt;a href="https://prometheus.io/docs/instrumenting/exposition_formats/">Prometheus&lt;/a>/&lt;a href="https://github.com/OpenObservability/OpenMetrics">OpenMetrics&lt;/a> exposition format, including support for Windows 10 via &lt;a href="https://github.com/prometheus-community/windows_exporter">windows_exporter&lt;/a>. Netdata will autodetect &lt;a href="https://github.com/netdata/go.d.plugin/blob/master/config/go.d/prometheus.conf">over 600 Prometheus endpoints&lt;/a> and instantly generate charts with all the exposed metrics, meaningfully visualized.
&lt;p>You can also quickly and easily configure the collector with the names and URLs of additional Prometheus endpoints to instantly view automatically generated charts with all the exposed metrics, meaningfully visualized within Netdata at the same high-granularity, per-second frequency you expect, all in real time. To learn more about how to configure, check out our &lt;a title="documentation" href="https://learn.netdata.cloud/docs/agent/collectors/go.d.plugin/modules/prometheus" target="_blank" rel="noopener noreferrer">documentation&lt;/a>.&lt;/p></description></item><item><title>Introducing the all-new Netdata Cloud</title><link>https://www.netdata.cloud/blog/introducing-the-all-new-netdata-cloud/</link><pubDate>Wed, 29 Jul 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/introducing-the-all-new-netdata-cloud/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16672" src="../wp-archive/uploads/2022/03/All-New-Cloud-1200x712.png" alt="" width="1200" height="712" />
&lt;p>In case you missed it, we released an all-new version of Netdata Cloud in May. &lt;a title="Netdata Cloud" href="https://staging-www.netdata.cloud/cloud/">Netdata Cloud&lt;/a> is a free service that can be accessed from any browser and provides you with a consolidated view of your entire infrastructure.&lt;/p>
&lt;p>Netdata Cloud works differently from other monitoring solutions. Most solutions limit the number and frequency of metrics because they rely on architectures that aggregate data. Netdata Cloud, however, streams limited metadata from each node running the Netdata Agent, keeping you in control of the data on your systems. The advantage of this architecture is that there is &lt;strong>no limit&lt;/strong> to the number or frequency of metrics, regardless of the scale or complexity of your IT infrastructure. You can truly monitor every metric, from every system and application, across your entire infrastructure, in real time. For free.&lt;/p></description></item><item><title>Sysadmin Day 2020: IT Heroes and Homelabs</title><link>https://www.netdata.cloud/blog/sysadmin-day-2020-it-heroes-and-homelabs/</link><pubDate>Mon, 27 Jul 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/sysadmin-day-2020-it-heroes-and-homelabs/</guid><description>&lt;!--truncate-->
&lt;p dir="ltr" lang="en">&lt;img class="alignnone size-large wp-image-16685" src="../wp-archive/uploads/2022/03/Group-1-2-1200x825.png" alt="" width="1200" height="825" />&lt;/p>
&lt;p dir="ltr" lang="en">Sysadmin Day 2020 is right around the corner and we’d like to show our appreciation for all the sysadmins out there who keep IT humming along and come to the rescue to resolve critical issues day in and day out. This year, we’re celebrating all week long by hosting an IT Heroes and Homelabs contest. &lt;strong>Join the celebration by retweeting our post with the hashtag #SysadminDay #NetdataWin, and we’ll enter you in a drawing to win some Netdata swag!&lt;/strong>&lt;/p></description></item><item><title>The Netdata Community</title><link>https://www.netdata.cloud/blog/the-netdata-community/</link><pubDate>Mon, 27 Jul 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/the-netdata-community/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16700" src="../wp-archive/uploads/2022/03/Athens-company-meetup-1_2019-scaled-e1588700274782-1200x747.jpeg" alt="" width="1200" height="747" />
&lt;p>Netdata users and contributors comprise a large, global, but somewhat fragmented community – or a set of communities. You can find us on IRC (#netdata on freenode), on &lt;a href="https://www.reddit.com/r/netdata/">Reddit&lt;/a>, on &lt;a href="https://twitter.com/linuxnetdata">social media&lt;/a>, and, of course, on &lt;a href="https://github.com/netdata/netdata">GitHub&lt;/a>, where the main open-source Netdata project repo lives. And yes, you can find us on other platforms as well.&lt;/p>
&lt;p> &lt;/p>
&lt;p>GitHub is a great way to get in touch with the Netdata team and project contributors to tell them about bugs or to discuss new features. But we realized that bug reports are not a conversation that works for everyone. A more informal and easier-to-access communication channel would provide a better way for the community to congregate and engage, and would also provide an easier way for everybody to talk to the team. We wanted to provide the community a home.&lt;/p></description></item><item><title>Why Netdata picked VerneMQ</title><link>https://www.netdata.cloud/blog/why-netdata-picked-vernemq/</link><pubDate>Tue, 14 Jul 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/why-netdata-picked-vernemq/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-full wp-image-16705" src="../wp-archive/uploads/2022/03/Blog-Why-Netdata-Picked-VerneMQ.jpeg" alt="" width="683" height="470" />
&lt;p>In 2019, the Netdata team already knew that a Netdata Cloud solution in the form of an online platform would greatly complement Netdata’s distributed monitoring by making it much easier to organize large infrastructures and by enabling new ways for teams to collaborate. The old node registry available at the time wasn’t enough for Netdata’s users.&lt;/p>
&lt;p>Building an online platform, even one that does not directly process users’ metrics, is challenging. But less challenging than it was even a few years ago, since the technology stack has improved greatly over the years.&lt;/p></description></item><item><title>What is Infrastructure Monitoring?</title><link>https://www.netdata.cloud/blog/what-is-infrastructure-monitoring/</link><pubDate>Tue, 30 Jun 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/what-is-infrastructure-monitoring/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-medium wp-image-16343" src="../wp-archive/uploads/2022/03/Blog-What_is_Infrastructure_Monitoring_Header-600x450.png" alt="" width="600" height="450" />
&lt;p>IT is advancing blazingly fast. To keep up with architectural changes and hybrid environments, it’s more important than ever to maintain efficient infrastructure monitoring and troubleshooting. Adding to the complexity is the increase of distributed systems, comprised of many components and services. For IT teams to effectively manage monitoring modern infrastructure, it’s necessary to have the right practices and tools in place that enable teams to do their jobs as quickly as possible with fewer resources.&lt;/p></description></item><item><title>Release 1.23: Kubernetes &amp; eBPF Observability</title><link>https://www.netdata.cloud/blog/release-1-23/</link><pubDate>Thu, 25 Jun 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/release-1-23/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16712" src="../wp-archive/uploads/2022/03/Agent-1.23-release-1200x900.png" alt="" width="1200" height="900" />
&lt;p>As adoption of container infrastructure grows in popularity, we’re continuing to focus on the most effective ways users can quickly and easily deploy container monitoring to instantly get access to deep, real-time insights. Agent release 1.23 introduces service discovery for Kubernetes clusters, monitoring for individual nodes, and eBPF monitoring per application on an event frequency for quickly identifying the root cause. Quickly shed light on your infrastructure performance with these new features!&lt;/p></description></item><item><title>The role of shift-left testing in an agile environment</title><link>https://www.netdata.cloud/blog/agile-static-analysis/</link><pubDate>Tue, 14 Apr 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/agile-static-analysis/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16723" src="../wp-archive/uploads/2022/03/Netdata-Security-use-with-Static-Analysis-1200x899.png" alt="" width="1200" height="899" />
&lt;p>With the rapid growth of security threats to infrastructure, it’s more important than ever to proactively address vulnerabilities. As an open-source project, built on the trust of users and contributors, Netdata has security concerns at its core.&lt;/p>
&lt;p>Because we’re committed to code security and quality, we apply &lt;a href="https://agilemanifesto.org/">Agile principles&lt;/a> throughout the software development process. A component of this includes regular static analysis. Through continuous, automated testing, we’re able to move quickly to keep up with end-user requests without compromising our source code.&lt;/p></description></item><item><title>Release 1.21: New Collectors &amp; Faster Exporters</title><link>https://www.netdata.cloud/blog/release-1-21/</link><pubDate>Mon, 06 Apr 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/release-1-21/</guid><description>&lt;!--truncate-->
&lt;div class="et_pb_module et_pb_text et_pb_text_0 et_pb_text_align_left et_pb_bg_layout_light">
&lt;div class="et_pb_text_inner">
&lt;img class="alignnone size-full wp-image-16737" src="../wp-archive/uploads/2022/03/release-1.21.0.png" alt="" width="1200" height="600" />
&lt;p>We’re in the middle of a scary, uncertain time, and we hope those of you reading are staying safe and healthy.&lt;/p>
&lt;p>Despite the current challenges, the 40+ members of the &lt;a title="Netdata Remote Working" href="https://staging-www.netdata.cloud/blog/culture/netdata-remote-working/">remote-first Netdata&lt;/a> team have been hard at work on the next version of the Netdata Agent: v1.21.0.&lt;/p>
&lt;p>This release is foundational: While we do have fantastic new collectors and three new ways to export your metrics for long-term storage, many of the most significant changes aren’t even those you’ll notice. While they may be beneath the hood, they’re going to power some amazing new features, UX improvements, and design overhauls.&lt;/p></description></item><item><title>Creating A Thriving, Agile, Remote Team</title><link>https://www.netdata.cloud/blog/netdata-remote-working/</link><pubDate>Wed, 25 Mar 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-remote-working/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-large wp-image-16750" src="../wp-archive/uploads/2022/03/netdata-remote-working_01-1200x899.png" alt="" width="1200" height="899" />
&lt;p>The coronavirus (COVID-19) pandemic has forced many organizations to take unprecedented steps towards remote working. As a fully distributed team, we’ve faced the common challenges of remote work. Based on our experience from our very beginning in 2018, all but a few of these organizations new to remote working will face hurdles to overcome and may try to revert to colocation as soon as possible. Remote working is hard, even when it’s carefully planned and executed. When the transition is rushed and seen as a necessary, temporary inconvenience, challenges are all but inevitable.&lt;/p></description></item><item><title>The Netdata Culture and People</title><link>https://www.netdata.cloud/blog/netdata-culture-people/</link><pubDate>Mon, 23 Mar 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/netdata-culture-people/</guid><description>&lt;!--truncate-->
&lt;p>&lt;img src="../wp-archive/uploads/2022/03/people-culture_01.png" alt="">&lt;/p>
&lt;p>There are many things I absolutely love about Netdata, but I’m most proud of our people and culture. Some words about this unique experience are long overdue.&lt;/p>
&lt;p>In a career that spans over two decades and six other companies of various sizes, nothing compares to the satisfaction of working in a company like ours. My answer to the canned interview question, “Where do you see yourself in 5 years”, was always the same: I don’t care; I just want to be solving problems and working with good people, real professionals, who I can trust and respect. In retrospect, I was missing another huge part of the equation, which is to mention the kind of company I wanted to work for. “Culture eats strategy for breakfast” is a cliche. More importantly, bad culture devours people’s souls; it sucks out any creative energy one may have, reducing engagement and, therefore, productivity. Short-term wins at the expense of company culture guarantee huge losses in the long run.&lt;/p></description></item><item><title>Contribute to Netdata’s machine learning efforts!</title><link>https://www.netdata.cloud/blog/contribute-machine-learning/</link><pubDate>Mon, 16 Mar 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/contribute-machine-learning/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-full wp-image-16783" src="../wp-archive/uploads/2022/03/contribute-machine-learning.png" alt="" width="991" height="1072" />
&lt;p>Netdata contributors have greatly influenced the growth of our company and are essential to our success. The time and expertise that contributors volunteer are fundamental to our goal of helping you build extraordinary infrastructures. We highly value end-user feedback during product development, which is why we’re looking to involve you in progressing our machine learning (ML) efforts! &lt;span id="more-2975">&lt;/span>As we are continually looking for ways to improve and enhance Netdata, we are starting to explore how we can leverage machine learning to introduce new product features. Our main focus at the moment is around automated &lt;a href="https://en.wikipedia.org/wiki/Anomaly_detection">anomaly detection&lt;/a>. This is a really interesting and challenging problem (high volume, high dimensional data, lack of ground truth labels, and so on), but we should be able to use some of the metrics monitored by Netdata to deliver new, awesome product features and user experiences (AI is the &lt;a href="https://www.gsb.stanford.edu/insights/andrew-ng-why-ai-new-electricity">new electricity&lt;/a>, after all 😃). However, developing ML-driven product features is quite different than traditional software development (see steps 1 to 7 in the picture above). Mainly, this is because you never really know what specific data transformations, problem formulation, and sets of algorithms will work best in advance. (&lt;a href="https://www.kdnuggets.com/2019/09/no-free-lunch-data-science.html">Here&lt;/a> is a good article explaining things, and if you really want to go down a rabbit hole, check out this &lt;a href="https://ai.stackexchange.com/questions/15650/what-are-the-implications-of-the-no-free-lunch-theorem-for-machine-learning">Stack Overflow question&lt;/a> and this &lt;a href="https://www.quora.com/What-does-the-No-Free-Lunch-theorem-mean-for-machine-learning-In-what-ways-do-popular-ML-algorithms-overcome-the-limitations-set-by-this-theorem">Quora thread&lt;/a>). Ideally, you first need to prototype your solution “in the lab” on some data you have already collected and do a few iterations of data → problem formulation → prototype. This process gives you a level of confidence in what you are doing (and some data to back it up) to move on to the even-more-complicated step of going from prototype to production. At Netdata, we are currently trying to get to step 4, where we can first prototype some solutions on real-world data and come up with ways to measure progress.&lt;/p></description></item><item><title>Linux eBPF monitoring with Netdata</title><link>https://www.netdata.cloud/blog/linux-ebpf-monitoring-with-netdata/</link><pubDate>Fri, 21 Feb 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/linux-ebpf-monitoring-with-netdata/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-full wp-image-16799" src="../wp-archive/uploads/2022/03/linux-ebpf-monitoring-netdata.png" alt="" width="1200" height="600" />
&lt;p>Your application isn’t finished when you’ve closed the last &lt;code>if&lt;/code> block and you lined up all the brackets. There’s a whole other world of testing, debugging, and optimization that you haven’t even touched yet.&lt;/p>
&lt;p>To help you more safely step into that complex phase of making your application &lt;em>even better&lt;/em>, we’ve just released a brand-new eBPF collector in &lt;a href="https://staging-www.netdata.cloud/blog/product/release-1.20/">v1.20 of Netdata&lt;/a>. With this collector enabled, you can monitor real-time metrics of Linux kernel functions and actions from the very same monitoring and troubleshooting dashboard you use for watching entire systems, or even entire infrastructures.&lt;/p></description></item><item><title>Release 1.20: Kernel Monitoring &amp; Infra-Wide Labels</title><link>https://www.netdata.cloud/blog/release-1-20/</link><pubDate>Fri, 21 Feb 2020 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/release-1-20/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-full wp-image-16790" src="../wp-archive/uploads/2022/03/release-1.20.0.png" alt="" width="1200" height="600" />
&lt;p>In Netdata’s first major release of 2020, we’re introducing two new features on the opposite ends of the monitoring spectrum.&lt;/p>
&lt;p>On one hand, we’re releasing an eBPF collector, which lets you collect, monitor, and visualize incredibly precise metrics straight from the Linux kernel. On the other, we added the ability to label agents to help you organize entire infrastructures and see &lt;em>every&lt;/em> important piece of information about streaming nodes in one place.&lt;/p></description></item><item><title>Redefining monitoring with Netdata (and how it came to be)</title><link>https://www.netdata.cloud/blog/redefining-monitoring-with-netdata/</link><pubDate>Thu, 19 Dec 2019 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/redefining-monitoring-with-netdata/</guid><description>&lt;!--truncate-->
&lt;p>&lt;img src="../wp-archive/uploads/2019/12/redefining-monitoring-netdata_01.png" alt="">&lt;/p>
&lt;h2 id="how-netdata-was-born">How Netdata was born&lt;/h2>
&lt;p>In 2013, I worked for a company that relied on financial transactions. We had a very simple SLA: complete all financial transactions within 3 seconds.&lt;/p>
&lt;p>We were migrating the infrastructure from colocated (physical servers) to the cloud (VMs). The transition was not smooth. We had a lot of issues on the cloud side, which we couldn’t even detect. Business metrics were randomly reporting significant loss of volume and a very bad SLA, but at the operational level we saw no issues—everything seemed to be working perfectly. Traces were showing a large delay in several transactions, but there were no failures.&lt;/p></description></item><item><title>Release 1.19: Web Log Parsing &amp; Unit Testing</title><link>https://www.netdata.cloud/blog/release-1-19/</link><pubDate>Wed, 27 Nov 2019 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/release-1-19/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-full wp-image-16837" src="../wp-archive/uploads/2022/03/release-1.19.0.png" alt="" width="1200" height="600" />
&lt;p>Network monitoring is complex, which is why we’re developing a monitoring tool that will drastically increase DevOps productivity. This release is all about improving Netdata’s day-in, day-out performance. We’re working hard to make deploy enhancements that help engineers make faster, smarter decisions about their systems.&lt;/p>
&lt;p> &lt;/p>
&lt;p>v1.19 of Netdata delivers a vastly improved way to collect, parse, and understand the health and performance of any service or application that runs through an Apache or Nginx web server.&lt;/p></description></item><item><title>Agile Team Safety Harness With cmocka &amp; FOSS</title><link>https://www.netdata.cloud/blog/agile-team-cmocka-foss/</link><pubDate>Tue, 26 Nov 2019 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/agile-team-cmocka-foss/</guid><description>&lt;!--truncate-->
&lt;p>Netdata is made up from agile teams who are deeply committed to improving the usability of our product. We want to respond to our users and introduce in-demand features. Working directly with our community is the best way to make Netdata better.&lt;/p>
&lt;p> &lt;/p>
&lt;p>But we face the same the dilemma as all agile teams: &lt;strong>How do we do this safely?&lt;/strong>&lt;/p>
&lt;p>Safety means that we can move quickly without compromising the quality of our code. Because we want to move quickly, engage with our users’ desires, and keep quality high, we’re becoming very serious about adopting unit testing in our work.&lt;/p></description></item><item><title>Release 1.18: What’s new with the database engine?</title><link>https://www.netdata.cloud/blog/release-1-18/</link><pubDate>Sat, 19 Oct 2019 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/release-1-18/</guid><description>&lt;!--truncate-->
&lt;img class="alignnone size-full wp-image-16859" src="../wp-archive/uploads/2022/03/release-1.18.0.png" alt="" width="1200" height="600" />
&lt;p>As your infrastructure grows more complex, storing long-term metrics becomes difficult and costly to retain. Your team stars to limit the amount of historical data they archive, causing gaps in coverage. Anomalies start to slip through the cracks.&lt;/p>
&lt;p>Version 1.18 of Netdata aims to solve the monitoring metrics storage problem once and for all.&lt;/p>
&lt;p>Aside from 5 new collectors, 16 bug fixes, 27 improvements, and 20 documentation updates, here’s what you need to know.&lt;/p></description></item><item><title>Release 1.17: Collection frequency gets flexible</title><link>https://www.netdata.cloud/blog/release-1-17/</link><pubDate>Mon, 09 Sep 2019 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/release-1-17/</guid><description>&lt;!--truncate-->
&lt;p>The next version of Netdata has arrived! Aside from dozens of quality-of-life and papercut fixes, we’ve launched some new features we know you’ll be excited to use straight away.&lt;/p>
&lt;p>Let’s dive in.&lt;/p>
&lt;h2>What’s new?&lt;/h2>
Release v1.17.0 contains 38 bug fixes, 33 improvements, and 20 documentation updates.
&lt;p>You can, of course, view the full list at the &lt;a href="https://github.com/netdata/netdata/releases/tag/v1.17.0">v1.17.0 release notes&lt;/a> on GitHub. But, let’s talk details on a few of the improvements and changes most requested by the Netdata community.&lt;/p></description></item><item><title>How and why we’re bringing long-term storage to Netdata</title><link>https://www.netdata.cloud/blog/db-engine/</link><pubDate>Wed, 07 Aug 2019 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/db-engine/</guid><description>&lt;!--truncate-->
&lt;p>We’ve built a lot of amazing things into the open-source &lt;a href="https://github.com/netdata/netdata">Netdata&lt;/a> monitoring system. But, no matter how far we’ve come, we’ll always be proud of how little RAM it uses.&lt;/p>
&lt;p>Right now, Netdata stores metrics in your system’s RAM using a ridiculously efficient database. It only saves or loads historical metrics from disk when you restart it. With this system, Netdata can be both low-resource and exhaustive in its collection of real-time metrics.&lt;/p></description></item><item><title>Release 1.16.0: Smarter binaries and built-in TLS</title><link>https://www.netdata.cloud/blog/release-1-16/</link><pubDate>Fri, 19 Jul 2019 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/release-1-16/</guid><description>&lt;!--truncate-->
&lt;p>We’re excited to launch release v1.16.0 of the open-source &lt;a href="https://github.com/netdata/netdata/">Netdata monitoring agent&lt;/a>, which delivers real-time health monitoring and performance troubleshooting to nearly any system or application.&lt;/p>
&lt;p>This release also contains 40 bug fixes, 31 improvements, and 20 documentation updates—if you’d like to see the full list, check out the &lt;a href="https://github.com/netdata/netdata/releases/tag/v1.16.0">full release notes&lt;/a>.&lt;/p>
&lt;p>Details aside, I know people are going to be most curious about the big changes we’ve just delivered to Netdata—let’s dive in.&lt;/p></description></item><item><title>Open Source Contributions: Supporting The Community</title><link>https://www.netdata.cloud/blog/open-source-contributions/</link><pubDate>Tue, 02 Jul 2019 00:00:00 +0000</pubDate><guid>https://www.netdata.cloud/blog/open-source-contributions/</guid><description>&lt;!--truncate-->
&lt;p>Netdata &lt;em>must&lt;/em> be doing something right when it comes to inspiring contributions. Our &lt;a href="https://github.com/netdata/netdata">open-source, distributed monitoring agent&lt;/a> has &lt;img src="https://img.shields.io/github/stars/netdata/netdata.svg" alt="GitHub stars" /> on GitHub and has seen contributions from hundreds of people: &lt;img src="https://img.shields.io/github/contributors/netdata/netdata.svg" alt="GitHub contributors" />. We’ve even hired a handful of our contributors to work full-time on making the Netdata ecosystem even more powerful.&lt;/p>
&lt;p> &lt;/p>
&lt;p>The community is passionate about what we’re building, and they’re actively interested in making it work better for their particular needs.&lt;/p></description></item></channel></rss>