The only agent that thinks for itself

Autonomous Monitoring with self-learning AI built-in, operating independently across your entire stack.

Unlimited Metrics & Logs
Machine learning & MCP
5% CPU, 150MB RAM
3GB disk, >1 year retention
800+ integrations, zero config
Dashboards, alerts out of the box
> Discover Netdata Agents

Centralized metrics streaming and storage

Aggregate metrics from multiple agents into centralized Parent nodes for unified monitoring across your infrastructure.

Stream from unlimited agents
Long-term data retention
High availability clustering
Data replication & backup
Scalable architecture
Enterprise-grade security
> Learn about Parents

Fully managed cloud platform

Access your monitoring data from anywhere with our SaaS platform. No infrastructure to manage, automatic updates, and global availability.

Zero infrastructure management
99.9% uptime SLA
Global data centers
Automatic updates & patches
Enterprise SSO & RBAC
SOC2 & ISO certified
> Explore Netdata Cloud

Deploy Netdata Cloud in your infrastructure

Run the full Netdata Cloud platform on-premises for complete data sovereignty and compliance with your security policies.

Complete data sovereignty
Air-gapped deployment
Custom compliance controls
Private network integration
Dedicated support team
Kubernetes & Docker support
> Learn about Cloud On-Premises

Powerful, intuitive monitoring interface

Modern, responsive UI built for real-time troubleshooting with customizable dashboards and advanced visualization capabilities.

Real-time chart updates
Customizable dashboards
Dark & light themes
Advanced filtering & search
Responsive on all devices
Collaboration features
> Explore Netdata UI

Monitor on the go

Native iOS and Android apps bring full monitoring capabilities to your mobile device with real-time alerts and notifications.

iOS & Android apps
Push notifications
Touch-optimized interface
Offline data access
Biometric authentication
Widget support
> Download apps

The future of infrastructure observability

See our strategic direction across AI-native observability, full-stack signals, operational intelligence, and enterprise platform maturity.

AI-native observability
Full-stack signal coverage
Operational intelligence
Enterprise platform maturity
Agent releases every 6 weeks
Cloud continuous delivery
> Explore Product Roadmap

Best energy efficiency

True real-time per-second

100% automated zero config

Centralized observability

Multi-year retention

High availability built-in

Zero maintenance

Always up-to-date

Enterprise security

Complete data control

Air-gap ready

Compliance certified

Millisecond responsiveness

Infinite zoom & pan

Works on any device

Native performance

Instant alerts

Monitor anywhere

AI-native observability

Continuous delivery

Open source foundation

80% Faster Incident Resolution

AI-powered troubleshooting from detection, to root cause and blast radius identification, to reporting.

True Real-Time and Simple, even at Scale

Linearly and infinitely scalable full-stack observability, that can be deployed even mid-crisis.

90% Cost Reduction, Full Fidelity

Instead of centralizing the data, Netdata distributes the code, eliminating pipelines and complexity.

See and Map Your Entire Network

Live topology, flow analytics, and SNMP device and trap monitoring — unified with your full-stack observability.

Control Without Surrender

SOC 2 Type 2 certified with every metric kept on your infrastructure.

Integrations

800+ collectors and notification channels, auto-discovered and ready out of the box.

800+ data collectors
Auto-discovery & zero config
Cloud, infra, app protocols
Notifications out of the box
> Explore integrations
Real Results
46% Cost Reduction

Reduced monitoring costs by 46% while cutting staff overhead by 67%.

— Leonardo Antunez, Codyas

Zero Pipeline

No data shipping. No central storage costs. Query at the edge.

From Our Users
"Out-of-the-Box"

So many out-of-the-box features! I mostly don't have to develop anything.

— Simon Beginn, LANCOM Systems

No Query Language

Point-and-click troubleshooting. No PromQL, no LogQL, no learning curve.

Enterprise Ready
67% Less Staff, 46% Cost Cut

Enterprise efficiency without enterprise complexity—real ROI from day one.

— Leonardo Antunez, Codyas

SOC 2 Type 2 Certified

Zero data egress. Only metadata reaches the cloud. Your metrics stay on your infrastructure.

Full Coverage
800+ Collectors

Auto-discovered and configured. No manual setup required.

Any Notification Channel

Slack, PagerDuty, Teams, email, webhooks—all built-in.

Built for the People Who Get Paged

Because 3am alerts deserve instant answers, not hour-long hunts.

Every Industry Has Rules. We Master Them.

See how healthcare, finance, and government teams cut monitoring costs 90% while staying audit-ready.

Monitor Any Technology. Configure Nothing.

Install the agent. It already knows your stack.
From Our Users
"A Rare Unicorn"

Netdata gives more than you invest in it. A rare unicorn that obeys the Pareto rule.

— Eduard Porquet Mateu, TMB Barcelona

99% Downtime Reduction

Reduced website downtime by 99% and cloud bill by 30% using Netdata alerts.

— Falkland Islands Government

Real Savings
30% Cloud Cost Reduction

Optimized resource allocation based on Netdata alerts cut cloud spending by 30%.

— Falkland Islands Government

46% Cost Cut

Reduced monitoring staff by 67% while cutting operational costs by 46%.

— Codyas

Real Coverage
"Plugin for Everything"

Netdata has agent capacity or a plugin for everything, including Windows and Kubernetes.

— Eduard Porquet Mateu, TMB Barcelona

"Out-of-the-Box"

So many out-of-the-box features! I mostly don't have to develop anything.

— Simon Beginn, LANCOM Systems

Real Speed
Troubleshooting in 30 Seconds

From 2-3 minutes to 30 seconds—instant visibility into any node issue.

— Matthew Artist, Nodecraft

20% Downtime Reduction

20% less downtime and 40% budget optimization from out-of-the-box monitoring.

— Simon Beginn, LANCOM Systems

Pay per Node. Unlimited Everything Else.

One price per node. Unlimited metrics, logs, users, and retention. No per-GB surprises.

Free tier—forever
No metric limits or caps
Retention you control
Cancel anytime
> See pricing plans

What's Your Monitoring Really Costing You?

Most teams overpay by 40-60%. Let's find out why.

Expose hidden metric charges
Calculate tool consolidation
Customers report 30-67% savings
Results in under 60 seconds
> See what you're really paying

Your Infrastructure Is Unique. Let's Talk.

Because monitoring 10 nodes is different from monitoring 10,000.

On-prem & air-gapped deployment
Volume pricing & agreements
Architecture review for your scale
Compliance & security support
> Start a conversation

Monitoring That Sells Itself

Deploy in minutes. Impress clients in hours. Earn recurring revenue for years.

30-second live demos close deals
Zero config = zero support burden
Competitive margins & deal protection
Response in 48 hours
> Apply to partner

Per-Second Metrics at Homelab Prices

Same engine, same dashboards, same ML. Just priced for tinkerers.

Community: Free forever · 5 nodes · non-commercial
Homelab: $90/yr · unlimited nodes · fair usage
> Get the Homelab Plan

$1,000 Per Referral. Unlimited Referrals.

Your colleagues get 10% off. You get 10% commission. Everyone wins.

10% of subscriptions, up to $1,000 each
Track earnings inside Netdata Cloud
PayPal/Venmo payouts in 3-4 weeks
No caps, no complexity
> Get your referral link
Cost Proof
40% Budget Optimization

"Netdata's significant positive impact" — LANCOM Systems

Calculate Your Savings

Compare vs Datadog, Grafana, Dynatrace

Savings Proof
46% Cost Reduction

"Cut costs by 46%, staff by 67%" — Codyas

30% Cloud Bill Savings

"Reduced cloud bill by 30%" — Falkland Islands Gov

Enterprise Proof
"Better Than Combined Alternatives"

"Better observability with Netdata than combining other tools." — TMB Barcelona

Real Engineers, <24h Response

DPA, SLAs, on-prem, volume pricing

Why Partners Win
Demo Live Infrastructure

One command, 30 seconds, real data—no sandbox needed

Zero Tickets, High Margins

Auto-config + per-node pricing = predictable profit

Homelab Ready
Free Video Course

8-episode Netdata tutorial by LearnLinux.tv

76k+ GitHub Stars

3rd most starred monitoring project

Worth Recommending
Product That Delivers

Customers report 40-67% cost cuts, 99% downtime reduction

Zero Risk to Your Rep

Free tier lets them try before they buy

AI Support Assistant, Available 24/7

Nedi has access to all official documentation, source code, and resources. Ask any question about Netdata—responds in your language.

Deployment & configuration
Troubleshooting & sizing
Alerts & notifications
Evidence-based answers
> Ask Nedi now

Never Fight Fires Alone

Docs, community, and expert help—pick your path to resolution.

Learn.netdata.cloud docs
Discord, Forums, GitHub
Premium support available
> Get answers now

60 Seconds to First Dashboard

One command to install. Zero config. 850+ integrations documented.

Linux, Windows, K8s, Docker
Auto-discovers your stack
> Read our documentation

76,000+ Engineers Strong

615+ contributors. 1.5M daily downloads. One mission: simplify observability.

Per-Second. 90% Cheaper. Data Stays Home.

Side-by-side comparisons: costs, real-time granularity, and data sovereignty for every major tool.

See why teams switch from Datadog, Prometheus, Grafana, and more.

> Browse all comparisons
Edge-Native Observability, Born Open Source
Per-second visibility, ML on every metric, and data that never leaves your infrastructure.
Founded in 2016
615+ contributors worldwide
Remote-first, engineering-driven
Open source first
> Read our story
Promises We Publish—and Prove
12 principles backed by open code, independent validation, and measurable outcomes.
Open source, peer-reviewed
Zero config, instant value
Data sovereignty by design
Aligned pricing, no surprises
> See all 12 principles
Edge-Native, AI-Ready, 100% Open
76k+ stars. Full ML, AI, and automation—GPLv3+, not premium add-ons.
76,000+ GitHub stars
GPLv3+ licensed forever
ML on every metric, included
Zero vendor lock-in
> Explore our open source
Build Real-Time Observability for the World
Remote-first team shipping per-second monitoring with ML on every metric.
Remote-first, fully distributed
Open source (76k+ stars)
Challenging technical problems
Your code on millions of systems
> See open roles
Meet the Team Behind Netdata
Conferences, meetups, and tradeshows where you can see Netdata in action and talk to the engineers who build it.
Live demos and deep dives
Book 1-on-1 meetings
Talks and panel sessions
Event recaps and photos
> See all events
Talk to a Netdata Human in <24 Hours
Sales, partnerships, press, or professional services—real engineers, fast answers.
Discuss your observability needs
Pricing and volume discounts
Partnership opportunities
Media and press inquiries
> Book a conversation
Your Data. Your Rules.
On-prem data, cloud control plane, transparent terms.
Trust & Scale
76,000+ GitHub Stars

One of the most popular open-source monitoring projects

SOC 2 Type 2 Certified

Enterprise-grade security and compliance

Data Sovereignty

Your metrics stay on your infrastructure

Validated
University of Amsterdam

"Most energy-efficient monitoring solution" — ICSOC 2023, peer-reviewed

ADASTEC (Autonomous Driving)

"Doesn't miss alerts—mission-critical trust for safety software"

Community Stats
615+ Contributors

Global community improving monitoring for everyone

1.5M+ Downloads/Day

Trusted by teams worldwide

GPLv3+ Licensed

Free forever, fully open source agent

Why Join?
Remote-First

Work from anywhere, async-friendly culture

Impact at Scale

Your work helps millions of systems

Databases

What Is Cardinality In Databases: A Comprehensive Guide

Why Cardinality Is Key To Database Performance
by Netdata Team · January 29, 2024

An Introduction To Cardinality

Cardinality is a fundamental concept in databases that plays a crucial role in designing efficient databases and optimizing query performance. For those new to databases, understanding what is cardinality is essential for effective data management. This comprehensive guide will explain the meaning of cardinality, its types, and its impact on database performance, making it accessible for both beginners and intermediate users.

What Is Cardinality In Databases?

Cardinality in databases refers to the uniqueness of data values contained in a column. It essentially measures how many distinct values exist in a column compared to the total number of rows in a table.

Cardinality can be understood in two main ways:

  • Mathematical Sense: The number of elements in a set.
  • Database Context: The number of unique values in a column, which helps in optimizing how data is stored and retrieved.

The Importance Of Cardinality In Databases

Understanding cardinality is vital for database performance and efficiency. It affects query optimization, indexing, and overall database design.

How Cardinality Impacts Query Optimization

Database query optimizers use cardinality to determine the most efficient way to execute queries. Knowing how many unique values are in a column helps the optimizer choose the best method to retrieve data.

For example, in an e-commerce database, the ProductID column typically has high cardinality because each product has a unique ID. This makes it ideal for indexing. On the other hand, the Category column might have low or medium cardinality because many products share the same category.

The Role Of Cardinality In Indexing

  • High Cardinality: Columns with high cardinality are great for indexing because they allow the database to quickly locate specific rows.
  • Low Cardinality: Columns with low cardinality are less effective for indexing as they result in larger sets of data to be scanned.

Real-World Use Cases Of Cardinality In Databases

Cardinality has a direct impact on how well a database performs in different real-world scenarios. Here are a few examples of how it plays a role across various industries:

E-commerce

In a product database, the ProductID column usually has high cardinality, as each item is unique. Indexing this column ensures fast lookup and smooth checkout experiences. Conversely, the Category or Brand columns tend to have lower cardinality and are less useful for indexing.

Finance

Transaction logs often include high-cardinality fields such as TransactionID or Timestamp, which are essential for fraud detection and audit trails. Mismanaging these columns can lead to bloated indexes and slow queries.

Healthcare

In patient records, fields like PatientID and VisitDate have high cardinality. Efficient indexing of these columns is crucial for retrieving patient histories quickly, especially in emergency scenarios.

SaaS Analytics

Multi-tenant platforms often track user behavior across applications. Fields like SessionID or UserID can have extremely high cardinality. Monitoring and managing these effectively is key to maintaining performance.

By understanding how cardinality affects common workloads, teams can make better decisions when designing schemas, building indexes, or writing queries.

High Cardinality vs Low Cardinality: What’s The Difference?

High Cardinality

High cardinality refers to columns with many unique values. These columns are typically used for primary keys or unique identifiers. Example: In a user database, the Email column would have high cardinality because each user has a unique email address. This uniqueness makes it suitable for indexing, allowing for fast searches and data retrieval.

Low Cardinality

Low cardinality refers to columns with few unique values. These columns are often used for categorical data. Example: In a survey database, the Gender column would have low cardinality with values like “Male” and “Female”. Since there are only a few distinct values, indexing this column might not significantly speed up queries.

Cardinality In SQL Databases

In SQL databases, cardinality affects query execution plans. When you execute a query, the database engine uses cardinality estimates to determine the most efficient way to retrieve data.

How Cardinality Affects Query Performance

  • Execution Plans: The query optimizer creates execution plans based on cardinality to minimize resource usage.
  • Statistics: Databases maintain statistics about cardinality, often stored as histograms, to help the optimizer make accurate decisions.

Example Of SQL Cardinality

Consider the following SQL query:

SELECT * FROM employees WHERE department_id = 5;

If the department_id column has low cardinality (few departments), the optimizer might choose a full table scan. However, if it has high cardinality (many departments), it might use an index to quickly find the matching rows.

Cardinality vs Selectivity: Key Differences

Cardinality

As mentioned above, cardinality refers to the number of distinct or unique values present in a database column. High cardinality indicates many unique values, while low cardinality means there are fewer distinct values with many repetitions.

For example:

In a “Customer ID” column, each customer might have a unique ID, leading to high cardinality. In a “Country” column for an international company, there may be fewer distinct values (e.g., USA, UK, India), resulting in low cardinality. Cardinality helps databases decide how to process queries by giving an understanding of the data distribution within a column.

Selectivity

Selectivity, on the other hand, refers to the fraction of rows that a database query will return based on a condition applied to a column. It is a ratio between the number of matching rows and the total number of rows in the table.

Selectivity is expressed as a value between 0 and 1, where:

  • A selectivity of 1 means all rows match the query (low selectivity).
  • A selectivity of 0 means no rows match the query.
  • A selectivity close to 0 means the query returns a very small fraction of rows (high selectivity).

For example:

If a “Customer ID” column has a query condition like WHERE CustomerID = 123, only one row will likely match because the column is highly unique, resulting in high selectivity. If a query is applied to a “Gender” column with a condition WHERE Gender = ‘Female’, and half the database consists of females, this would result in low selectivity (because many rows match the condition).

How Cardinality And Selectivity Work Together

While cardinality measures the uniqueness of values in a column, selectivity measures how “exclusive” a query condition is in returning rows. Higher selectivity generally leads to more efficient queries, as fewer rows are returned, whereas lower selectivity can indicate a broader query that returns many rows.

Both concepts are important for database query optimization:

  • High cardinality columns (with many unique values) typically offer high selectivity when queried, making them good candidates for indexing.
  • Low cardinality columns (with many repeated values) often have low selectivity, which can make indexes less effective for query performance.

Monitoring & Managing Cardinality

Effectively managing cardinality involves using database management tools to monitor and analyze data distribution.

Database Management Tools & Techniques

  • Monitoring Tools: Tools like Netdata Database Performance Monitor can help track and understand cardinality in your database.
  • Statistics Updates: Regularly update statistics to ensure the optimizer has accurate information for query planning.

Cardinality Challenges In Time-Series Monitoring Systems

In time-series databases used for observability and monitoring, such as Prometheus or InfluxDB, cardinality can become a major challenge.

High-cardinality metrics, typically caused by excessive or overly granular labels, can lead to:

Increased Storage Usage

Every unique combination of labels creates a new time series, multiplying the amount of data stored.

Slower Queries

As the number of series grows, queries take longer to execute, especially when scanning across many time series.

Label Explosion

This happens when dynamic values like user IDs, hostnames, or timestamps are used as metric labels, resulting in millions of unique series.

For example, a metric like http_requests_total{user_id=““123"”} repeated for every user can generate enormous cardinality, putting strain on the system.

Best Practice: Use labels thoughtfully. Avoid dynamic values as label keys, and periodically audit your metrics to control label cardinality.

Monitoring tools like Netdata offer real-time visibility into cardinality trends, helping you detect and fix high-cardinality issues before they affect performance.

3 Best Practices For Working With Cardinality

1. Regular Monitoring

Keep an eye on how cardinality changes over time.

2. Optimize Indexes

Adjust indexes based on cardinality to improve performance.

3. Update Statistics

Ensure database statistics are up-to-date for accurate query optimization.

Common Pitfalls When Managing Cardinality

While understanding cardinality is important, applying it incorrectly can lead to inefficiencies and performance bottlenecks. Here are some common mistakes to avoid:

Over-Indexing Low-Cardinality Columns

Creating indexes on columns with very few unique values (e.g., Gender, Country) may increase storage usage without improving query performance.

Underestimating High-Cardinality Dimensions In Analytics

In business intelligence systems, visualizing high-cardinality fields (like CustomerEmail) can overwhelm dashboards and slow down queries.

Ignoring Data Growth

A column might start with low cardinality but grow over time. Failing to monitor these changes can cause outdated statistics and suboptimal query plans.

Label Explosion In Observability Tools

Tagging metrics with too many high-cardinality labels in monitoring platforms can lead to excessive memory usage and degraded performance.

Avoiding these anti-patterns can help ensure your database and observability stack remain performant and scalable over time.

Final Thoughts On Cardinality In Databases

Understanding what is cardinality in databases is crucial for designing efficient databases and optimizing query performance. By knowing the types of cardinality and their impact on database operations, you can make informed decisions about indexing, query optimization, and overall data management.

For further insights and tools, explore additional resources on database performance and optimization to enhance your database management skills. By mastering the concept of cardinality, you’ll be better equipped to manage and optimize your databases effectively

Cardinality In Databases - Frequently Asked Questions (FAQs)

What Is Considered High Cardinality?

High cardinality refers to columns or metrics that have a large number of unique values, such as Email, TransactionID, or UserID.

Is High Cardinality Good Or Bad?

It depends. High cardinality is useful for indexing and precise querying but can increase memory usage and slow performance if not managed correctly.

How Can I Find Cardinality Issues In My Database?

Use your database’s query planner, histograms, or monitoring tools like Netdata to inspect data distribution and identify high or low cardinality columns.

How Do PostgreSQL, MySQL, Or Oracle Handle Cardinality?

These databases maintain statistics on column cardinality to help their query optimizers choose the most efficient execution plans.

Does High Cardinality Affect Dashboards Or Analytics?

Yes. High-cardinality dimensions in BI tools can slow down dashboards and overwhelm filters or aggregations.