The only agent that thinks for itself

Autonomous Monitoring with self-learning AI built-in, operating independently across your entire stack.

Unlimited Metrics & Logs
Machine learning & MCP
5% CPU, 150MB RAM
3GB disk, >1 year retention
800+ integrations, zero config
Dashboards, alerts out of the box
> Discover Netdata Agents

Centralized metrics streaming and storage

Aggregate metrics from multiple agents into centralized Parent nodes for unified monitoring across your infrastructure.

Stream from unlimited agents
Long-term data retention
High availability clustering
Data replication & backup
Scalable architecture
Enterprise-grade security
> Learn about Parents

Fully managed cloud platform

Access your monitoring data from anywhere with our SaaS platform. No infrastructure to manage, automatic updates, and global availability.

Zero infrastructure management
99.9% uptime SLA
Global data centers
Automatic updates & patches
Enterprise SSO & RBAC
SOC2 & ISO certified
> Explore Netdata Cloud

Deploy Netdata Cloud in your infrastructure

Run the full Netdata Cloud platform on-premises for complete data sovereignty and compliance with your security policies.

Complete data sovereignty
Air-gapped deployment
Custom compliance controls
Private network integration
Dedicated support team
Kubernetes & Docker support
> Learn about Cloud On-Premises

Powerful, intuitive monitoring interface

Modern, responsive UI built for real-time troubleshooting with customizable dashboards and advanced visualization capabilities.

Real-time chart updates
Customizable dashboards
Dark & light themes
Advanced filtering & search
Responsive on all devices
Collaboration features
> Explore Netdata UI

Monitor on the go

Native iOS and Android apps bring full monitoring capabilities to your mobile device with real-time alerts and notifications.

iOS & Android apps
Push notifications
Touch-optimized interface
Offline data access
Biometric authentication
Widget support
> Download apps

The future of infrastructure observability

See our strategic direction across AI-native observability, full-stack signals, operational intelligence, and enterprise platform maturity.

AI-native observability
Full-stack signal coverage
Operational intelligence
Enterprise platform maturity
Agent releases every 6 weeks
Cloud continuous delivery
> Explore Product Roadmap

Best energy efficiency

True real-time per-second

100% automated zero config

Centralized observability

Multi-year retention

High availability built-in

Zero maintenance

Always up-to-date

Enterprise security

Complete data control

Air-gap ready

Compliance certified

Millisecond responsiveness

Infinite zoom & pan

Works on any device

Native performance

Instant alerts

Monitor anywhere

AI-native observability

Continuous delivery

Open source foundation

80% Faster Incident Resolution

AI-powered troubleshooting from detection, to root cause and blast radius identification, to reporting.

True Real-Time and Simple, even at Scale

Linearly and infinitely scalable full-stack observability, that can be deployed even mid-crisis.

90% Cost Reduction, Full Fidelity

Instead of centralizing the data, Netdata distributes the code, eliminating pipelines and complexity.

See and Map Your Entire Network

Live topology, flow analytics, and SNMP device and trap monitoring — unified with your full-stack observability.

Control Without Surrender

SOC 2 Type 2 certified with every metric kept on your infrastructure.

Integrations

800+ collectors and notification channels, auto-discovered and ready out of the box.

800+ data collectors
Auto-discovery & zero config
Cloud, infra, app protocols
Notifications out of the box
> Explore integrations
Real Results
46% Cost Reduction

Reduced monitoring costs by 46% while cutting staff overhead by 67%.

— Leonardo Antunez, Codyas

Zero Pipeline

No data shipping. No central storage costs. Query at the edge.

From Our Users
"Out-of-the-Box"

So many out-of-the-box features! I mostly don't have to develop anything.

— Simon Beginn, LANCOM Systems

No Query Language

Point-and-click troubleshooting. No PromQL, no LogQL, no learning curve.

Enterprise Ready
67% Less Staff, 46% Cost Cut

Enterprise efficiency without enterprise complexity—real ROI from day one.

— Leonardo Antunez, Codyas

SOC 2 Type 2 Certified

Zero data egress. Only metadata reaches the cloud. Your metrics stay on your infrastructure.

Full Coverage
800+ Collectors

Auto-discovered and configured. No manual setup required.

Any Notification Channel

Slack, PagerDuty, Teams, email, webhooks—all built-in.

Built for the People Who Get Paged

Because 3am alerts deserve instant answers, not hour-long hunts.

Every Industry Has Rules. We Master Them.

See how healthcare, finance, and government teams cut monitoring costs 90% while staying audit-ready.

Monitor Any Technology. Configure Nothing.

Install the agent. It already knows your stack.
From Our Users
"A Rare Unicorn"

Netdata gives more than you invest in it. A rare unicorn that obeys the Pareto rule.

— Eduard Porquet Mateu, TMB Barcelona

99% Downtime Reduction

Reduced website downtime by 99% and cloud bill by 30% using Netdata alerts.

— Falkland Islands Government

Real Savings
30% Cloud Cost Reduction

Optimized resource allocation based on Netdata alerts cut cloud spending by 30%.

— Falkland Islands Government

46% Cost Cut

Reduced monitoring staff by 67% while cutting operational costs by 46%.

— Codyas

Real Coverage
"Plugin for Everything"

Netdata has agent capacity or a plugin for everything, including Windows and Kubernetes.

— Eduard Porquet Mateu, TMB Barcelona

"Out-of-the-Box"

So many out-of-the-box features! I mostly don't have to develop anything.

— Simon Beginn, LANCOM Systems

Real Speed
Troubleshooting in 30 Seconds

From 2-3 minutes to 30 seconds—instant visibility into any node issue.

— Matthew Artist, Nodecraft

20% Downtime Reduction

20% less downtime and 40% budget optimization from out-of-the-box monitoring.

— Simon Beginn, LANCOM Systems

Pay per Node. Unlimited Everything Else.

One price per node. Unlimited metrics, logs, users, and retention. No per-GB surprises.

Free tier—forever
No metric limits or caps
Retention you control
Cancel anytime
> See pricing plans

What's Your Monitoring Really Costing You?

Most teams overpay by 40-60%. Let's find out why.

Expose hidden metric charges
Calculate tool consolidation
Customers report 30-67% savings
Results in under 60 seconds
> See what you're really paying

Your Infrastructure Is Unique. Let's Talk.

Because monitoring 10 nodes is different from monitoring 10,000.

On-prem & air-gapped deployment
Volume pricing & agreements
Architecture review for your scale
Compliance & security support
> Start a conversation

Monitoring That Sells Itself

Deploy in minutes. Impress clients in hours. Earn recurring revenue for years.

30-second live demos close deals
Zero config = zero support burden
Competitive margins & deal protection
Response in 48 hours
> Apply to partner

Per-Second Metrics at Homelab Prices

Same engine, same dashboards, same ML. Just priced for tinkerers.

Community: Free forever · 5 nodes · non-commercial
Homelab: $90/yr · unlimited nodes · fair usage
> Get the Homelab Plan

$1,000 Per Referral. Unlimited Referrals.

Your colleagues get 10% off. You get 10% commission. Everyone wins.

10% of subscriptions, up to $1,000 each
Track earnings inside Netdata Cloud
PayPal/Venmo payouts in 3-4 weeks
No caps, no complexity
> Get your referral link
Cost Proof
40% Budget Optimization

"Netdata's significant positive impact" — LANCOM Systems

Calculate Your Savings

Compare vs Datadog, Grafana, Dynatrace

Savings Proof
46% Cost Reduction

"Cut costs by 46%, staff by 67%" — Codyas

30% Cloud Bill Savings

"Reduced cloud bill by 30%" — Falkland Islands Gov

Enterprise Proof
"Better Than Combined Alternatives"

"Better observability with Netdata than combining other tools." — TMB Barcelona

Real Engineers, <24h Response

DPA, SLAs, on-prem, volume pricing

Why Partners Win
Demo Live Infrastructure

One command, 30 seconds, real data—no sandbox needed

Zero Tickets, High Margins

Auto-config + per-node pricing = predictable profit

Homelab Ready
Free Video Course

8-episode Netdata tutorial by LearnLinux.tv

76k+ GitHub Stars

3rd most starred monitoring project

Worth Recommending
Product That Delivers

Customers report 40-67% cost cuts, 99% downtime reduction

Zero Risk to Your Rep

Free tier lets them try before they buy

AI Support Assistant, Available 24/7

Nedi has access to all official documentation, source code, and resources. Ask any question about Netdata—responds in your language.

Deployment & configuration
Troubleshooting & sizing
Alerts & notifications
Evidence-based answers
> Ask Nedi now

Never Fight Fires Alone

Docs, community, and expert help—pick your path to resolution.

Learn.netdata.cloud docs
Discord, Forums, GitHub
Premium support available
> Get answers now

60 Seconds to First Dashboard

One command to install. Zero config. 850+ integrations documented.

Linux, Windows, K8s, Docker
Auto-discovers your stack
> Read our documentation

76,000+ Engineers Strong

615+ contributors. 1.5M daily downloads. One mission: simplify observability.

Per-Second. 90% Cheaper. Data Stays Home.

Side-by-side comparisons: costs, real-time granularity, and data sovereignty for every major tool.

See why teams switch from Datadog, Prometheus, Grafana, and more.

> Browse all comparisons
Edge-Native Observability, Born Open Source
Per-second visibility, ML on every metric, and data that never leaves your infrastructure.
Founded in 2016
615+ contributors worldwide
Remote-first, engineering-driven
Open source first
> Read our story
Promises We Publish—and Prove
12 principles backed by open code, independent validation, and measurable outcomes.
Open source, peer-reviewed
Zero config, instant value
Data sovereignty by design
Aligned pricing, no surprises
> See all 12 principles
Edge-Native, AI-Ready, 100% Open
76k+ stars. Full ML, AI, and automation—GPLv3+, not premium add-ons.
76,000+ GitHub stars
GPLv3+ licensed forever
ML on every metric, included
Zero vendor lock-in
> Explore our open source
Build Real-Time Observability for the World
Remote-first team shipping per-second monitoring with ML on every metric.
Remote-first, fully distributed
Open source (76k+ stars)
Challenging technical problems
Your code on millions of systems
> See open roles
Meet the Team Behind Netdata
Conferences, meetups, and tradeshows where you can see Netdata in action and talk to the engineers who build it.
Live demos and deep dives
Book 1-on-1 meetings
Talks and panel sessions
Event recaps and photos
> See all events
Talk to a Netdata Human in <24 Hours
Sales, partnerships, press, or professional services—real engineers, fast answers.
Discuss your observability needs
Pricing and volume discounts
Partnership opportunities
Media and press inquiries
> Book a conversation
Your Data. Your Rules.
On-prem data, cloud control plane, transparent terms.
Trust & Scale
76,000+ GitHub Stars

One of the most popular open-source monitoring projects

SOC 2 Type 2 Certified

Enterprise-grade security and compliance

Data Sovereignty

Your metrics stay on your infrastructure

Validated
University of Amsterdam

"Most energy-efficient monitoring solution" — ICSOC 2023, peer-reviewed

ADASTEC (Autonomous Driving)

"Doesn't miss alerts—mission-critical trust for safety software"

Community Stats
615+ Contributors

Global community improving monitoring for everyone

1.5M+ Downloads/Day

Trusted by teams worldwide

GPLv3+ Licensed

Free forever, fully open source agent

Why Join?
Remote-First

Work from anywhere, async-friendly culture

Impact at Scale

Your work helps millions of systems

$ guides / smartctl-disk-monitoring / smartctl-g-sense-error-rate ▌

Operations Guides

G-Sense_Error_Rate rising: shock and vibration reaching the drive

G-Sense_Error_Rate (SMART attribute ID 221 on most drives, ID 191 on some) is a cumulative counter of shock and vibration events that exceeded the drive’s internal threshold, as detected by its built-in accelerometer. The attribute is HDD-only. SSDs and NVMe drives do not report it. Some enterprise HDD models do not expose the attribute at all, which does not mean vibration is absent, only that the drive does not instrument it.

When this number climbs, the first question is not “is the drive failing?” but “is the vibration causing damage?” A high raw value with no correlated media errors means the accelerometer is registering environmental noise. A growing raw value alongside rising Seek_Error_Rate (ID 7) or Reallocated_Sector_Ct (ID 5) means the vibration is causing the read/write head to contact the platter surface (head slap), producing permanent media damage.

The second question is “what is the source?” The accelerometer cannot tell you whether the vibration comes from the drive next door, a failing HVAC system, a rigid mounting tray, or construction in the building. It only tells you the drive felt something.

What this means

Smartmontools reports G-Sense_Error_Rate as a raw48 value under either attribute ID 221 or ID 191, depending on the drive. Two columns in smartctl -A output matter:

  • RAW_VALUE: the cumulative count of detected events since manufacture.
  • VALUE: the normalized score the firmware computes from the raw data. On most drives this starts at 100 or 200 and trends toward THRESH (typically 000) as accumulated events reach levels the firmware considers harmful.

The raw value alone is not actionable without context. Some drives, particularly Seagate IronWolf and Exos enterprise HDDs, have accelerometers sensitive enough to trigger on normal chassis vibration from fans and adjacent drives. Raw values in the hundreds or low thousands are common on these drives in standard server chassis, with the normalized value steady at 100 and no media errors. That is sensor noise, not damage.

What is actionable:

  • Any growth from the last reading: investigate the physical environment. TICKET.
  • Rapid growth alongside rising Seek_Error_Rate or Reallocated_Sector_Ct: the vibration is causing head slap. PAGE.

The normalized value declining from its starting point (typically 100) toward the threshold is the firmware’s own assessment that accumulated shock events are exceeding the drive’s design tolerance. A declining normalized value is always significant, regardless of the raw value.

flowchart TD
    A["G-Sense_Error_Rate
growing"] --> B{"Normalized VALUE
declining?"} B -- "No, stable at 100/200" --> C{"Seek_Error_Rate (ID 7)
or Reallocated_Sector_Ct (ID 5)
also rising?"} C -- "No" --> D["Likely benign sensor noise
TICKET: investigate environment"] C -- "Yes" --> E["Vibration causing
positioning errors or head slap"] B -- "Yes, declining" --> E E --> F["PAGE: fix vibration source
and plan drive replacement"] D --> G["Fix mounting, resonance,
or environmental vibration"]

If smartctl -A does not list G-Sense_Error_Rate under either ID, the drive either lacks an accelerometer or does not expose it through SMART. Absence of the attribute means you have no telemetry for it.

Common causes

CauseWhat it looks likeFirst thing to check
Dense-shelf resonanceMultiple drives in the same JBOD shelf show G-Sense growth simultaneously; seek errors may also riseCheck whether drives in the same physical shelf share the pattern
Missing anti-vibration mountsSingle drive shows growth after installation or relocation; no other drives affectedInspect drive mounting hardware (grommets, rubber washers, sled rails)
Environmental vibration (HVAC, construction, nearby machinery)Growth correlates with specific times of day or facility eventsAsk facilities about HVAC cycling, construction, or equipment changes
Shipping damageBrand-new drive with large raw value on first read; normalized value may already be degradedCompare against baseline; RMA if normalized is degraded on a new drive
Seagate sensor sensitivityHigh raw value (hundreds to thousands) but normalized stays at 100; no seek errors or reallocationsVerify normalized value and check for correlated damage signals

Quick checks

# Read the G-Sense_Error_Rate attribute (both possible IDs)
smartctl -A /dev/sdX | grep -i "g.sense"

# Check for correlated seek errors (vibration degrading head positioning)
smartctl -A /dev/sdX | grep -i "seek_error"

# Check for reallocated sectors (head slap damage)
smartctl -A /dev/sdX | grep -i "reallocat"

# Check for pending sectors (vibration-induced read failures)
smartctl -A /dev/sdX | grep -i "current_pending"

# Check whether other drives in the same chassis show similar patterns
for disk in /dev/sd?; do echo "=== $disk ==="; smartctl -A "$disk" 2>/dev/null | grep -i "g.sense"; done

# Review ATA error log for mechanical error signatures
smartctl -l error /dev/sdX | grep -iE "AMNF|CCTO|ABRT"

# Run a conveyance test to check for shipping/handling damage (~5 minutes, HDD only)
smartctl -t conveyance /dev/sdX

All commands except the conveyance test are read-only. The conveyance test runs a short surface scan designed to detect transport-related damage. It is safe to run on a production drive but will briefly compete with host I/O.

smartctl does print CCTO in ATA error-log records, but only for commands that use that error-bit mapping; it means Command Completion Timed Out, not a mechanical shock signature. The vibration-correlated signatures in this guide are AMNF, ABRT, and UNC.

How to diagnose it

  1. Record the raw and normalized values. The raw value is cumulative since manufacture. What matters is whether it changed since the last collection, and whether the normalized value has moved from its starting point (typically 100 or 200).

  2. Check for correlated damage signals. Pull Seek_Error_Rate (ID 7), Reallocated_Sector_Ct (ID 5), and Current_Pending_Sector (ID 197). If any are growing, the vibration is past noise and into active damage. Escalate from TICKET to PAGE.

  3. Scope the problem across the chassis. Run G-Sense checks on every drive in the same physical enclosure. If multiple drives show growth simultaneously, the root cause is environmental (shelf resonance, HVAC, power supply vibration). If only one drive is affected, the problem is likely its mounting.

  4. Compare against the deployment baseline. A new drive with a high raw value but normalized at 100 and no media errors likely accumulated sensor events during shipping. A drive that was zero and is now growing has an active vibration source.

  5. Identify the vibration source. This is physical investigation. Check drive mounting hardware, chassis fan health, nearby equipment, and facility conditions. The accelerometer tells you the drive felt something, not what caused it.

  6. Check the ATA error log for mechanical signatures. AMNF (Address Mark Not Found) errors indicate low-level servo or format problems that vibration can worsen. ABRT (Aborted Command) errors can indicate the drive aborted due to mechanical difficulty. If either appears alongside G-Sense growth, the vibration is actively degrading the drive.

Metrics and signals to monitor

SignalWhy it mattersWarning sign
G-Sense_Error_Rate raw value (ID 191/221)Cumulative count of vibration events exceeding the drive’s thresholdAny growth from last reading
G-Sense_Error_Rate normalized valueFirmware’s assessment of accumulated shock severityDecline from starting value (typically 100 or 200)
Seek_Error_Rate normalized value (ID 7)Vibration directly degrades head positioning accuracyNormalized value declining alongside G-Sense growth
Reallocated_Sector_Ct raw value (ID 5)Head slap from vibration damages platter surfaceAny increase
Current_Pending_Sector raw value (ID 197)Vibration-induced read failures mark sectors as suspectAny non-zero value
G-Sense_Error_Rate across chassisMultiple drives growing together indicates environmental sourceTwo or more drives in same shelf showing simultaneous growth
Drive temperature (ID 194)Failed fans increase chassis vibration and temperatureTemperature rising alongside G-Sense growth

Fixes

Fix the environment, not the drive. If the media is healthy (no seek errors, no reallocations, normalized G-Sense stable), replacing the drive will not solve the problem. The new drive will show the same growth because the vibration source is still present.

Dense-shelf resonance

In JBOD enclosures with many bays, adjacent spinning drives create sympathetic vibration that feeds back through the chassis:

  • Install anti-vibration grommets or rubber washers between the drive sled and the chassis.
  • Verify that drive rails or sleds are designed for vibration damping, not rigid metal-to-metal contact.
  • If the enclosure supports staggered spin-up, enable it to reduce transient vibration during power cycles.
  • Reduce drive density if the enclosure is overpopulated for its vibration isolation design.

Missing or degraded mounting hardware

A single drive showing G-Sense growth while neighbors are stable suggests a mounting problem. Inspect:

  • Anti-vibration grommets: present, not perished, correctly seated.
  • Drive sled rails: properly engaged, not loose.
  • Screws: torqued to spec. Over-tightening can compress rubber isolation solid, defeating its purpose.

Environmental vibration

If growth correlates with facility events (HVAC cycling, construction, nearby equipment):

  • Relocate the server or storage shelf away from the vibration source.
  • Install vibration isolation pads under the rack.
  • Coordinate with facilities to identify and mitigate the source.

Shipping damage on a new drive

A brand-new drive with a large G-Sense raw value and a degraded normalized value likely experienced shock during shipping. If the drive also shows media errors, file an RMA. Do not deploy a shipping-damaged drive.

When to replace the drive

Replace the drive only when vibration has caused confirmed media damage: Reallocated_Sector_Ct is growing, or Seek_Error_Rate normalized is declining alongside G-Sense growth. In that case:

  1. Fix the vibration source first (mounting, resonance, environment).
  2. Replace the damaged drive.
  3. Monitor the replacement for G-Sense growth. If the new drive shows the same pattern, the environment was not fixed.

Prevention

  • Baseline at deployment. Capture a full SMART snapshot when each drive is installed. The G-Sense raw value at deployment tells you what accumulated during shipping and factory testing. Alert on growth from this baseline, not on absolute value.

  • Use vibration-damped mounting hardware in every HDD in every chassis. Rigid metal-to-metal contact transmits chassis vibration directly to the drive.

  • Monitor G-Sense across the chassis. A single drive showing growth is a mounting problem. Multiple drives growing is an environmental problem. The distinction determines whether you fix a sled or call facilities.

  • Alert on growth and normalized value, not raw thresholds. A drive that sits at raw value 500 with normalized 100 and zero seek errors for its entire service life is healthy.

  • Run conveyance tests on newly deployed drives before committing them to production.

How Netdata helps

  • Rate of change detection. Netdata tracks G-Sense_Error_Rate changes over time, distinguishing a counter that jumped 200 in an hour from one that grew 200 over six months.

  • Correlation across signals. G-Sense_Error_Rate growth appears alongside Seek_Error_Rate changes and Reallocated_Sector_Ct in a single view, the exact correlation that separates benign sensor noise from head slap.

  • Cross-drive comparison. When multiple drives in the same host show G-Sense growth simultaneously, Netdata surfaces the chassis-wide pattern, pointing to an environmental root cause rather than a drive-specific one.

  • Anomaly detection flags G-Sense growth that deviates from each drive’s established baseline, reducing false alerts from drives with naturally high but stable raw values.