The only agent that thinks for itself

Autonomous Monitoring with self-learning AI built-in, operating independently across your entire stack.

Unlimited Metrics & Logs
Machine learning & MCP
5% CPU, 150MB RAM
3GB disk, >1 year retention
800+ integrations, zero config
Dashboards, alerts out of the box
> Discover Netdata Agents

Centralized metrics streaming and storage

Aggregate metrics from multiple agents into centralized Parent nodes for unified monitoring across your infrastructure.

Stream from unlimited agents
Long-term data retention
High availability clustering
Data replication & backup
Scalable architecture
Enterprise-grade security
> Learn about Parents

Fully managed cloud platform

Access your monitoring data from anywhere with our SaaS platform. No infrastructure to manage, automatic updates, and global availability.

Zero infrastructure management
99.9% uptime SLA
Global data centers
Automatic updates & patches
Enterprise SSO & RBAC
SOC2 & ISO certified
> Explore Netdata Cloud

Deploy Netdata Cloud in your infrastructure

Run the full Netdata Cloud platform on-premises for complete data sovereignty and compliance with your security policies.

Complete data sovereignty
Air-gapped deployment
Custom compliance controls
Private network integration
Dedicated support team
Kubernetes & Docker support
> Learn about Cloud On-Premises

Powerful, intuitive monitoring interface

Modern, responsive UI built for real-time troubleshooting with customizable dashboards and advanced visualization capabilities.

Real-time chart updates
Customizable dashboards
Dark & light themes
Advanced filtering & search
Responsive on all devices
Collaboration features
> Explore Netdata UI

Monitor on the go

Native iOS and Android apps bring full monitoring capabilities to your mobile device with real-time alerts and notifications.

iOS & Android apps
Push notifications
Touch-optimized interface
Offline data access
Biometric authentication
Widget support
> Download apps

The future of infrastructure observability

See our strategic direction across AI-native observability, full-stack signals, operational intelligence, and enterprise platform maturity.

AI-native observability
Full-stack signal coverage
Operational intelligence
Enterprise platform maturity
Agent releases every 6 weeks
Cloud continuous delivery
> Explore Product Roadmap

Best energy efficiency

True real-time per-second

100% automated zero config

Centralized observability

Multi-year retention

High availability built-in

Zero maintenance

Always up-to-date

Enterprise security

Complete data control

Air-gap ready

Compliance certified

Millisecond responsiveness

Infinite zoom & pan

Works on any device

Native performance

Instant alerts

Monitor anywhere

AI-native observability

Continuous delivery

Open source foundation

80% Faster Incident Resolution

AI-powered troubleshooting from detection, to root cause and blast radius identification, to reporting.

True Real-Time and Simple, even at Scale

Linearly and infinitely scalable full-stack observability, that can be deployed even mid-crisis.

90% Cost Reduction, Full Fidelity

Instead of centralizing the data, Netdata distributes the code, eliminating pipelines and complexity.

See and Map Your Entire Network

Live topology, flow analytics, and SNMP device and trap monitoring — unified with your full-stack observability.

Control Without Surrender

SOC 2 Type 2 certified with every metric kept on your infrastructure.

Integrations

800+ collectors and notification channels, auto-discovered and ready out of the box.

800+ data collectors
Auto-discovery & zero config
Cloud, infra, app protocols
Notifications out of the box
> Explore integrations
Real Results
46% Cost Reduction

Reduced monitoring costs by 46% while cutting staff overhead by 67%.

— Leonardo Antunez, Codyas

Zero Pipeline

No data shipping. No central storage costs. Query at the edge.

From Our Users
"Out-of-the-Box"

So many out-of-the-box features! I mostly don't have to develop anything.

— Simon Beginn, LANCOM Systems

No Query Language

Point-and-click troubleshooting. No PromQL, no LogQL, no learning curve.

Enterprise Ready
67% Less Staff, 46% Cost Cut

Enterprise efficiency without enterprise complexity—real ROI from day one.

— Leonardo Antunez, Codyas

SOC 2 Type 2 Certified

Zero data egress. Only metadata reaches the cloud. Your metrics stay on your infrastructure.

Full Coverage
800+ Collectors

Auto-discovered and configured. No manual setup required.

Any Notification Channel

Slack, PagerDuty, Teams, email, webhooks—all built-in.

Built for the People Who Get Paged

Because 3am alerts deserve instant answers, not hour-long hunts.

Every Industry Has Rules. We Master Them.

See how healthcare, finance, and government teams cut monitoring costs 90% while staying audit-ready.

Monitor Any Technology. Configure Nothing.

Install the agent. It already knows your stack.
From Our Users
"A Rare Unicorn"

Netdata gives more than you invest in it. A rare unicorn that obeys the Pareto rule.

— Eduard Porquet Mateu, TMB Barcelona

99% Downtime Reduction

Reduced website downtime by 99% and cloud bill by 30% using Netdata alerts.

— Falkland Islands Government

Real Savings
30% Cloud Cost Reduction

Optimized resource allocation based on Netdata alerts cut cloud spending by 30%.

— Falkland Islands Government

46% Cost Cut

Reduced monitoring staff by 67% while cutting operational costs by 46%.

— Codyas

Real Coverage
"Plugin for Everything"

Netdata has agent capacity or a plugin for everything, including Windows and Kubernetes.

— Eduard Porquet Mateu, TMB Barcelona

"Out-of-the-Box"

So many out-of-the-box features! I mostly don't have to develop anything.

— Simon Beginn, LANCOM Systems

Real Speed
Troubleshooting in 30 Seconds

From 2-3 minutes to 30 seconds—instant visibility into any node issue.

— Matthew Artist, Nodecraft

20% Downtime Reduction

20% less downtime and 40% budget optimization from out-of-the-box monitoring.

— Simon Beginn, LANCOM Systems

Pay per Node. Unlimited Everything Else.

One price per node. Unlimited metrics, logs, users, and retention. No per-GB surprises.

Free tier—forever
No metric limits or caps
Retention you control
Cancel anytime
> See pricing plans

What's Your Monitoring Really Costing You?

Most teams overpay by 40-60%. Let's find out why.

Expose hidden metric charges
Calculate tool consolidation
Customers report 30-67% savings
Results in under 60 seconds
> See what you're really paying

Your Infrastructure Is Unique. Let's Talk.

Because monitoring 10 nodes is different from monitoring 10,000.

On-prem & air-gapped deployment
Volume pricing & agreements
Architecture review for your scale
Compliance & security support
> Start a conversation

Monitoring That Sells Itself

Deploy in minutes. Impress clients in hours. Earn recurring revenue for years.

30-second live demos close deals
Zero config = zero support burden
Competitive margins & deal protection
Response in 48 hours
> Apply to partner

Per-Second Metrics at Homelab Prices

Same engine, same dashboards, same ML. Just priced for tinkerers.

Community: Free forever · 5 nodes · non-commercial
Homelab: $90/yr · unlimited nodes · fair usage
> Get the Homelab Plan

$1,000 Per Referral. Unlimited Referrals.

Your colleagues get 10% off. You get 10% commission. Everyone wins.

10% of subscriptions, up to $1,000 each
Track earnings inside Netdata Cloud
PayPal/Venmo payouts in 3-4 weeks
No caps, no complexity
> Get your referral link
Cost Proof
40% Budget Optimization

"Netdata's significant positive impact" — LANCOM Systems

Calculate Your Savings

Compare vs Datadog, Grafana, Dynatrace

Savings Proof
46% Cost Reduction

"Cut costs by 46%, staff by 67%" — Codyas

30% Cloud Bill Savings

"Reduced cloud bill by 30%" — Falkland Islands Gov

Enterprise Proof
"Better Than Combined Alternatives"

"Better observability with Netdata than combining other tools." — TMB Barcelona

Real Engineers, <24h Response

DPA, SLAs, on-prem, volume pricing

Why Partners Win
Demo Live Infrastructure

One command, 30 seconds, real data—no sandbox needed

Zero Tickets, High Margins

Auto-config + per-node pricing = predictable profit

Homelab Ready
Free Video Course

8-episode Netdata tutorial by LearnLinux.tv

76k+ GitHub Stars

3rd most starred monitoring project

Worth Recommending
Product That Delivers

Customers report 40-67% cost cuts, 99% downtime reduction

Zero Risk to Your Rep

Free tier lets them try before they buy

AI Support Assistant, Available 24/7

Nedi has access to all official documentation, source code, and resources. Ask any question about Netdata—responds in your language.

Deployment & configuration
Troubleshooting & sizing
Alerts & notifications
Evidence-based answers
> Ask Nedi now

Never Fight Fires Alone

Docs, community, and expert help—pick your path to resolution.

Learn.netdata.cloud docs
Discord, Forums, GitHub
Premium support available
> Get answers now

60 Seconds to First Dashboard

One command to install. Zero config. 850+ integrations documented.

Linux, Windows, K8s, Docker
Auto-discovers your stack
> Read our documentation

76,000+ Engineers Strong

615+ contributors. 1.5M daily downloads. One mission: simplify observability.

Per-Second. 90% Cheaper. Data Stays Home.

Side-by-side comparisons: costs, real-time granularity, and data sovereignty for every major tool.

See why teams switch from Datadog, Prometheus, Grafana, and more.

> Browse all comparisons
Edge-Native Observability, Born Open Source
Per-second visibility, ML on every metric, and data that never leaves your infrastructure.
Founded in 2016
615+ contributors worldwide
Remote-first, engineering-driven
Open source first
> Read our story
Promises We Publish—and Prove
12 principles backed by open code, independent validation, and measurable outcomes.
Open source, peer-reviewed
Zero config, instant value
Data sovereignty by design
Aligned pricing, no surprises
> See all 12 principles
Edge-Native, AI-Ready, 100% Open
76k+ stars. Full ML, AI, and automation—GPLv3+, not premium add-ons.
76,000+ GitHub stars
GPLv3+ licensed forever
ML on every metric, included
Zero vendor lock-in
> Explore our open source
Build Real-Time Observability for the World
Remote-first team shipping per-second monitoring with ML on every metric.
Remote-first, fully distributed
Open source (76k+ stars)
Challenging technical problems
Your code on millions of systems
> See open roles
Meet the Team Behind Netdata
Conferences, meetups, and tradeshows where you can see Netdata in action and talk to the engineers who build it.
Live demos and deep dives
Book 1-on-1 meetings
Talks and panel sessions
Event recaps and photos
> See all events
Talk to a Netdata Human in <24 Hours
Sales, partnerships, press, or professional services—real engineers, fast answers.
Discuss your observability needs
Pricing and volume discounts
Partnership opportunities
Media and press inquiries
> Book a conversation
Your Data. Your Rules.
On-prem data, cloud control plane, transparent terms.
Trust & Scale
76,000+ GitHub Stars

One of the most popular open-source monitoring projects

SOC 2 Type 2 Certified

Enterprise-grade security and compliance

Data Sovereignty

Your metrics stay on your infrastructure

Validated
University of Amsterdam

"Most energy-efficient monitoring solution" — ICSOC 2023, peer-reviewed

ADASTEC (Autonomous Driving)

"Doesn't miss alerts—mission-critical trust for safety software"

Community Stats
615+ Contributors

Global community improving monitoring for everyone

1.5M+ Downloads/Day

Trusted by teams worldwide

GPLv3+ Licensed

Free forever, fully open source agent

Why Join?
Remote-First

Work from anywhere, async-friendly culture

Impact at Scale

Your work helps millions of systems

Buyer’s Guide - August 2026

The 9 best NATS monitoring tools, ranked

NATS is not a generic service you can point any APM agent at. Server stats, JetStream streams and consumers, routes, gateways, and leaf nodes all live behind NATS’s own HTTP monitoring API and system-account event stream, and only NATS-aware collectors exploit them. We ranked nine tools on metric depth, collection resolution, alerting, and total cost for production NATS fleets.

The 9 best NATS monitoring tools, ranked product interface

Why this list exists

NATS monitoring is a distinct problem from general infrastructure monitoring. nats-server exposes a rich HTTP monitoring API on port 8222 (varz, connz, routez, subsz, gatewayz, leafz, healthz, jsz, accountz) plus a system-account event stream. A collector that does not understand those endpoints gives you a fraction of the picture. The most common buying mistake is assuming a general-purpose APM or a default Prometheus scrape delivers NATS depth. In practice, most general tools stop at varz-level counters, while JetStream stream and consumer health, per-account usage, and gateway and leaf-node topology need a collector built for NATS semantics.

Three dimensions decide most outcomes:

  1. Metric depth. Does the tool cover more than basic varz counters? JetStream streams and consumers, per-account metrics, routes, gateways, leaf nodes, and connection-level pending bytes are what separate a real NATS monitor from a checkbox integration.
  2. Collection resolution. NATS is a high-throughput messaging system. Message-rate spikes and slow-consumer buildup happen in seconds, so a 10 to 15 second polling interval routinely misses the incident you are trying to explain.
  3. Time to value and cost shape. Auto-detection and prebuilt content beat a stack you wire together yourself, and per-host pricing punishes the many-small-nodes deployment pattern NATS fleets tend to follow.

A note on pricing: we do not quote competitor list prices in this guide. List prices for usage-based platforms are almost meaningless without knowing your ingestion volume, retention, and seat count, and quoting them out of context misleads more than it informs. Instead, each card describes the pricing shape and what makes the bill grow, with a link to the vendor’s own pricing page.

If you are still building your runbooks, our NATS monitoring guides walk through the operational side: which endpoints matter, how to read slow-consumer signals, and how to structure alerts for JetStream.

Methodology

How we evaluated NATS monitoring tools

We assembled the shortlist from tools with verified, documented NATS support: official NATS project utilities, dedicated collectors, and commercial platforms with a named NATS integration. Vendors with only a passing mention of NATS in a generic integrations catalog were excluded.

Weighting reflects what actually decides a NATS monitoring purchase. Metric depth carries the most weight because shallow coverage is the failure mode buyers hit first. Collection resolution is second because NATS incidents are fast. Alerting, time to value, and cost efficiency share the remainder, with ecosystem integration as a tiebreaker.

Tester credit

Compiled by the Netdata team - Updated August 12, 2026

Scoring criteria

  • NATS metric depth 25%
    JetStream, accounts, routes, gateways, leaf nodes, connection detail
  • Collection resolution 20%
    Per-second vs 10-15s polling for fast-moving message rates
  • Time to value 15%
    Auto-detection and prebuilt content vs assembled stacks
  • Alerting and anomaly detection 15%
    NATS-specific conditions, slow consumers, JetStream API errors
  • Cost efficiency 15%
    Pricing shape vs fleet size; operational cost of self-hosting
  • Ecosystem integration 10%
    Prometheus compatibility, Grafana, Kubernetes, alert routing

Vendor 01 / 09 · #netdata

01

Netdata

An open-source, per-second monitoring platform with a dedicated NATS collector covering server, JetStream, accounts, routes, gateways, and leaf nodes.

Netdata metrics tab showing real-time per-second charts of infrastructure metrics, representative of the NATS server, JetStream, and connection charts the NATS collector produces.

Best for

  • Teams running NATS across many servers or edge nodes who want per-second visibility without per-GB ingestion costs
  • Operators who want zero-config auto-detection of NATS instances on localhost or in Docker
  • Platform teams that want NATS metrics alongside host-level CPU, memory, and network context in one dashboard

Pricing

  • Per-node pricing: Netdata Cloud Business starts at $4.5/node/month on annual plans, with the per-node price decreasing as node count grows
  • Agents are open source (AGPL) and free to run
  • Free Cloud tier exists for small fleets
  • Bill grows with monitored node count, not with metrics volume or retention

Pros

  • Dedicated NATS collector (go.d.plugin nats module) with a 1-second default collection interval
  • Covers server traffic, messages, connections, CPU/memory/uptime, health probe status, JetStream streams, consumers, storage, API requests and errors, per-account metrics, routes, gateways, and leaf-node connections with RTT
  • Auto-detects NATS on localhost port 8222 and inside Docker containers with zero configuration
  • Supports multiple instances including remote servers, HTTP auth, TLS, and proxies
  • Machine-learning anomaly detection on collected metrics with no manual threshold tuning
  • 800+ integrations total, so NATS charts sit next to host, container, and application data

Where teams pair it

  • No alert rules ship preconfigured for the NATS collector; you create alerts for slow consumers and JetStream API errors yourself
  • No dedicated prebuilt NATS dashboard; you compose dashboards from the collected NATS charts
  • No NATS client application tracing; teams needing end-to-end request tracing pair it with an APM tool

Verdict

Netdata leads this list because it is the only option that combines the deepest open-source NATS metric coverage with a 1-second collection interval in a single package. Server, JetStream, account, route, gateway, and leaf-node scopes are all collected, and auto-detection on port 8222 means the agent finds NATS without configuration. The per-node pricing model keeps cost predictable as fleets grow across many small nodes, where per-host and per-GB models punish you. The honest gaps: no prebuilt NATS alert rules, no dedicated NATS dashboard out of the box, and the NATS HTTP monitoring port must be enabled, a prerequisite shared by every HTTP-based collector on this list.

Vendor 02 / 09 · #prometheus-grafana

02

Prometheus + Grafana

The de-facto open-source stack for NATS, pairing the official prometheus-nats-exporter with Prometheus storage and Grafana dashboards.

Best for

  • Teams already standardized on Prometheus and Grafana who want NATS in existing dashboards
  • Organizations that prefer assembling best-of-breed open-source components
  • Kubernetes-centric deployments where Prometheus is already the cluster standard

Pricing

  • Open source and self-hosted: you run and operate Prometheus, Grafana, and the exporter yourself
  • Grafana Cloud offers usage-based SaaS tiers with free and paid plans
  • Cost grows with metric cardinality, retention, and the operational effort of running the stack

Pros

  • Official prometheus-nats-exporter maintained by nats-io (Apache-2.0, active as of July 2026) aggregates varz, connz, subz, routez, healthz, and jsz endpoints
  • Exporter supports detailed connection metrics, account stats, gateway, leaf, and JetStream stream/consumer metrics via flags
  • Community Grafana dashboards exist for NATS servers and JetStream (dashboard IDs 2279 and 14725)
  • PromQL gives flexible alerting on any NATS metric
  • Huge ecosystem: Alertmanager, Thanos, Grafana Cloud, and hundreds of compatible tools

Cons

  • Requires assembling and operating exporter, Prometheus, and Grafana as separate components
  • Default scrape intervals of 10-15s miss sub-second NATS traffic spikes unless tuned
  • JetStream metrics create high cardinality in Prometheus; operators must filter jsz metrics carefully
  • No NATS-specific anomaly detection; alerting means writing manual PromQL thresholds

Verdict

This is the standard community answer for NATS monitoring, and for good reason: the official exporter covers nearly every NATS monitoring endpoint, and the ecosystem around it is unmatched. It ranks below Netdata because you are assembling three components, and default scrape intervals are too coarse for NATS traffic patterns. If your organization already runs Prometheus well, this is the path of least resistance. If you do not, budget real operational time for it. Open source here means self-hosted: the license is free, the operations are not.

Vendor 03 / 09 · #nats-surveyor

03

NATS Surveyor

A NATS-native Prometheus exporter from the NATS team that observes an entire deployment from a single system-account connection.

Best for

  • NATS operators who want one exporter to see the whole cluster without per-server sidecars
  • Teams on NATS 2.0+ system accounts who want per-account, per-route, and JetStream visibility
  • Synadia platform users who want a lighter open-source alternative to Synadia Insights

Pricing

  • Open source (Apache-2.0) and self-hosted: you operate Surveyor plus Prometheus and Grafana
  • Docker Compose stack included for quick start
  • Cost is operational: running the exporter, Prometheus storage, and Grafana

Pros

  • Maintained by nats-io (Apache-2.0, active as of August 2026); used extensively by Synadia
  • Single exporter connects to one NATS server and discovers the entire deployment via system-account Statz messages
  • Per-account metrics, gateway metrics, Raft group metrics, and JetStream stream/consumer metrics with leader-only and filter options to control cardinality
  • Ships with a docker-compose stack (Surveyor + Prometheus + Grafana) and a prebuilt dashboard
  • Provides a nats_up metric to distinguish exporter connectivity problems from NATS problems

Cons

  • Requires NATS system account credentials and system accounts enabled on the server
  • Requires Prometheus and Grafana (or another Prometheus-compatible stack) for storage and visualization
  • JetStream metrics can cause high cardinality in Prometheus unless jsz filters and leader-only mode are used
  • No built-in alerting or anomaly detection; relies on Alertmanager rules

Verdict

Surveyor is the most NATS-native open-source exporter, designed specifically for whole-deployment visibility through the system account. One connection gives you per-account, gateway, Raft, and JetStream data without an exporter per server, which scales better than prometheus-nats-exporter on large clusters. It ranks third because it is still only the collection layer: without Prometheus and Grafana behind it, it shows you nothing. For NATS 2.0+ deployments with system accounts already in place, it is the strongest exporter choice.

Vendor 04 / 09 · #synadia-insights

04

Synadia Insights

A commercial, NATS-native observability product from the company behind NATS that indexes everything a system account can observe into a queryable time-series graph.

Best for

  • Production deployments where NATS is critical infrastructure and teams need entity-level diagnosis (which connection, which account, which stream)
  • Teams that want 100+ built-in NATS checks with severity and remediation guidance without writing their own rules
  • Organizations that want AI-agent queryable NATS state without granting agents direct system access

Pricing

  • Annual contract priced per number of NATS servers monitored; a significant commitment for smaller teams
  • 30-day free trial with a built-in simulator for evaluation
  • Runs as a single read-only binary on a machine of your choosing; nothing installed on NATS nodes

Pros

  • Purpose-built for NATS by Synadia, the company behind the NATS project
  • Indexes server state, cluster topology, route health, connection metadata, subject-level message counts, JetStream asset state, account usage, and entity relationships
  • 100+ built-in automated checks with severity, plain-language description, and remediation guidance
  • Read-only design: cannot edit configurations, cannot access message payloads or stream data
  • Ships with an AI agent skill so agents can query NATS state in plain English
  • Complements existing Grafana, Datadog, and OpenTelemetry stacks rather than replacing them

Cons

  • Requires NATS Server v2.10+ and system account credentials
  • Annual contract pricing is a significant commitment for smaller teams
  • Does not replace a general monitoring stack; positioned as a NATS-specific complement
  • Does not work with Synadia Cloud, since it requires direct system account access

Verdict

This is the deepest NATS-specific diagnostics available commercially. It goes beyond metrics into entity relationships and configuration-change context, and the 100+ built-in checks are the opposite of Netdata’s and Prometheus’s write-your-own-alerts model. It ranks fourth because the annual contract and NATS v2.10+ requirement narrow the audience: for teams where NATS is critical infrastructure, it is compelling; for everyone else, it is more than they need. Note that it complements, rather than replaces, a general metrics platform.

Vendor 05 / 09 · #new-relic

05

New Relic

A commercial observability platform with a NATS quickstart that instruments NATS applications via the Go agent, with a prebuilt dashboard and starter alerts.

Best for

  • Teams already on New Relic who want NATS visibility alongside APM, logs, and traces in one platform
  • Organizations that want a prebuilt NATS dashboard and four starter alerts without building from scratch
  • Application teams that want client-side NATS visibility rather than server-side endpoint scraping

Pricing

  • Usage-based: per-user seats plus data ingestion volume
  • A free tier is available with ingestion limits for small-scale evaluation
  • Bill grows with ingestion volume, retention, and full-platform user count

Pros

  • Official NATS quickstart built by New Relic with one dashboard, four alerts, and documentation
  • Starter alerts cover memory above 90%, transaction errors above 10%, CPU above 90%, and Apdex below 0.5
  • Go agent instrumentation gives transaction tracing, error inbox, and service maps for NATS-based applications
  • Deep platform integration with logs, traces, and infrastructure monitoring

Cons

  • Quickstart is application-centric (Go agent) and does not scrape NATS HTTP monitoring endpoints for server or JetStream metrics
  • No JetStream stream or consumer metrics in the quickstart
  • Usage-based pricing can grow quickly with ingestion volume across a large NATS fleet
  • Requires embedding the New Relic Go agent in each NATS client application

Verdict

New Relic’s NATS support is a different thing from what most of this list offers: it watches your NATS client applications, not the NATS servers. That is genuinely useful for application teams who want tracing and service maps, and the quickstart is the fastest prebuilt-alert experience here. But for the operators reading this guide, the absence of server-side varz scraping and any JetStream coverage is disqualifying as a primary NATS monitor. Treat it as an APM complement, not a NATS monitoring solution.

Vendor 06 / 09 · #elastic

06

Elastic Observability

Elastic’s Metricbeat NATS module scrapes stats, connections, routes, subscriptions, and optional JetStream metrics into Elasticsearch and Kibana.

Best for

  • Teams already running Elasticsearch and Kibana who want NATS metrics alongside logs and APM
  • Organizations that want a single platform for NATS metrics, NATS logs, and application traces
  • Teams that prefer the Elastic Agent and Metricbeat unified collection model

Pricing

  • Usage-based SaaS (Elastic Cloud) priced on ingestion volume and retention
  • Open source self-managed option: you run and operate Elasticsearch, Kibana, and Metricbeat
  • Bill grows with data volume, not node count

Pros

  • Official Metricbeat NATS module with default metricsets: stats, connections, routes, subscriptions
  • Optional connection, route, and jetstream metricsets for per-stream and per-consumer detail
  • Tested with NATS 2.2.6 and 2.11.x
  • Ships with a predefined Kibana dashboard for NATS
  • Supports TLS, HTTP auth, and standard Metricbeat configuration options

Cons

  • Documented example uses a 10s collection period, coarser than per-second monitoring
  • JetStream metricset requires explicit per-account, per-stream, and per-consumer name configuration
  • Running Elasticsearch and Kibana is operationally heavy for NATS monitoring alone
  • No NATS-specific anomaly detection; relies on Elastic ML jobs or manual Kibana alerts

Verdict

The Metricbeat module is solid server-side coverage, and it is one of the few commercial integrations with a real JetStream metricset. Two things hold it back for NATS-first buyers: the 10s example period is too coarse for messaging traffic, and the JetStream metricset needs you to enumerate accounts, streams, and consumers by name, which does not scale with dynamic workloads. If your organization already runs Elastic, enabling the module is an easy win. Standing up the Elastic stack just for NATS is not.

Vendor 07 / 09 · #datadog

07

Datadog

Datadog’s NATS support is a community-maintained gnatsd check in integrations-extras collecting 25 server, connection, and route metrics.

Best for

  • Teams already standardized on Datadog who want basic NATS server visibility in existing dashboards
  • Organizations that accept community-tier integration support for NATS
  • Teams monitoring NATS alongside a large commercial SaaS infrastructure stack

Pricing

  • Per-host pricing plus per-GB log and metric ingestion
  • Community integration is free to install but requires a paid Datadog subscription
  • Bill grows with host count and ingestion volume

Pros

  • Collects 25 metrics across varz, connz, and routez: traffic, messages, connections, subscriptions, slow consumers, memory, routes, remotes
  • Service check gnatsd.can_connect reports CRITICAL when the Agent cannot reach the endpoint
  • Metrics tagged with cluster names for multi-cluster environments
  • Compatible with all major platforms

Cons

  • Community integration: not included in the Agent package, installed separately via datadog-agent integration install
  • No JetStream metrics at all; no account, gateway, or leaf-node coverage
  • No NATS-specific events or prebuilt dashboards
  • Community integrations get best-effort support rather than Datadog’s standard SLA

Verdict

Datadog’s NATS check is the weakest coverage-per-dollar on this list. Twenty-five varz, connz, and routez metrics is basic server health, and the complete absence of JetStream support is a serious gap for any modern NATS deployment where streams and consumers are the product. The community-tier status means best-effort support and a manual install step. If you are already deep in Datadog and only need to know the server is alive and how busy it is, it works. If JetStream matters to you, look elsewhere or pair it with a NATS-aware collector.

Vendor 08 / 09 · #instana

08

IBM Instana

IBM Instana provides automatic NATS monitoring with application discovery, service mapping, tracing, and automated troubleshooting inside its APM platform.

Best for

  • Enterprises already using Instana for APM who want NATS included in automatic discovery
  • Teams that value automatic instrumentation over manual exporter configuration
  • Organizations running NATS inside a larger microservices estate with tracing needs

Pricing

  • Per-host per-month pricing, billed annually across Essentials and Standard tiers
  • Self-hosted option priced per managed virtual server
  • Bill grows with the number of monitored hosts

Pros

  • NATS is a listed supported technology with automatic application discovery and service mapping
  • Provides tracing and automated troubleshooting for NATS-based applications
  • Agent-based automatic instrumentation requires no manual exporter setup
  • Part of a broad APM platform covering hundreds of technologies

Cons

  • NATS monitoring is application-centric (tracing and service maps) rather than server-infrastructure-centric
  • No documented JetStream stream or consumer metrics collection
  • Per-host pricing is expensive for large NATS fleets spread across many small nodes
  • NATS is one of hundreds of supported technologies; NATS-specific depth is limited compared to NATS-native tools

Verdict

Instana’s NATS story mirrors New Relic’s: it is about the applications talking to NATS, not the NATS servers themselves. Automatic discovery and service mapping are real conveniences, and if Instana is already your APM, NATS comes along for free in operational terms. But there is no documented JetStream metric collection, and per-host pricing stacks up painfully on the many-small-nodes topology NATS fleets favor. It ranks eighth because it answers a different question than the one this guide is about.

Vendor 09 / 09 · #nats-top

09

nats-top

A top-like terminal tool from the NATS team for live inspection of server stats and per-connection activity.

Best for

  • Operators who want an instant terminal view of a NATS server without installing a platform
  • Quick debugging: connection counts, message rates, slow consumers, per-connection stats
  • Scripted one-shot snapshots via the -o flag

Pricing

  • Open source (MIT) and self-hosted: build or download the binary and run it
  • No ongoing cost; operational footprint is a single CLI tool

Pros

  • Maintained by nats-io (MIT license, active as of June 2026)
  • Shows server load (CPU, memory, slow consumers), in/out message and byte rates, and per-connection detail including subscriptions, pending bytes, and uptime
  • Interactive sorting by connection ID, subscriptions, messages, bytes, or idle time
  • Supports TLS, basic auth, and HTTPS monitoring endpoints
  • Installable via curl script, go install, or RPM/Deb packages

Cons

  • Terminal-only: no dashboards, no history, no alerting
  • No JetStream, account, gateway, or leaf-node metrics
  • Single-server view; does not aggregate a cluster
  • A debugging utility, not a monitoring platform

Verdict

nats-top earns its place as the tool you reach for in the first sixty seconds of an incident: is the server loaded, which connection is backing up, are slow consumers accumulating. The 1-second refresh and per-connection pending bytes make it genuinely good at that job. But it keeps no history, fires no alerts, and sees one server at a time, so it cannot be your monitoring answer. Install it everywhere anyway; it pairs well with every platform on this list.

Frequently asked questions