Open-source utilities for tracking real-time resource utilization, hardware health, and system metrics across servers.
dockprom is a monitoring stack based on Prometheus and Grafana designed to track the performance of Docker containers and their underlying hosts. It functions as a complete solution for gathering real-time metrics and displaying them through a self-hosted dashboard. The project includes a suite of tools for collecting container and host metrics, as well as a discovery tool specifically for automatically identifying and adding tagged EC2 instances to the monitoring configuration. The system covers several observability areas, including time-series data storage and the creation of performance
Dockprom is a complete monitoring stack built on Prometheus and Grafana, providing real-time dashboards, alerting, and time-series storage for Docker hosts and containers, and it is fully self-hostable via Docker Compose.
Netdata is a real-time infrastructure monitoring tool and multi-node observability platform. It functions as a high-resolution monitoring agent, log and metric aggregator, and time-series database designed to provide full-stack visibility into server health. The system is distinguished by its per-second metric sampling and zero-configuration auto-discovery, which allows for immediate infrastructure tracking upon installation. It utilizes edge-based machine learning and unsupervised models to detect system anomalies and abnormal metric patterns locally on each node. For distributed environment
Netdata is a real-time infrastructure monitoring platform with per-second metric collection, built-in alerting, and interactive dashboards, and it supports multi-node setups and self-hosting, making it an excellent fit for this search.
Prometheus is a comprehensive monitoring and alerting platform designed to track infrastructure health and application performance. It functions as a time series database that ingests, indexes, and queries high-frequency numerical data points. By utilizing a pull-based model, the system periodically collects multi-dimensional metrics from monitored targets, storing them in an optimized block storage format that supports high-throughput ingestion and efficient historical analysis. The platform distinguishes itself through a specialized query engine that enables real-time analysis of performanc
Prometheus is a comprehensive monitoring and alerting platform with a built-in time-series database, pull-based metric collection via exporters, integrated alerting, and multi-server support, perfectly fitting the need for self-hosted server performance monitoring.
Netdata is a distributed observability platform designed for real-time infrastructure monitoring and performance tracking. It functions as a high-frequency agent that collects system, container, and application metrics with per-second precision, providing both local visualization and centralized aggregation across complex, multi-cloud environments. The platform distinguishes itself through edge-based intelligence, utilizing local machine learning models to automatically detect performance anomalies without requiring manual configuration or external query engines. Its architecture prioritizes
Netdata is a comprehensive, self-hostable distributed observability platform that collects server metrics via per-second agents, provides real-time dashboards and anomaly alerts, supports multi-server aggregation, and offers extensive plugin/agent extensibility — exactly the kind of tool this search is after.
SigNoz is a full-stack observability platform designed to collect, store, and visualize metrics, logs, and distributed traces in a unified environment. It leverages OpenTelemetry-based data collection to ingest telemetry from diverse sources using vendor-neutral protocols, ensuring interoperability across complex microservices architectures. The platform utilizes a high-performance columnar storage engine to enable rapid aggregation and filtering, providing a centralized backend for monitoring application health and performance. What distinguishes the platform is its focus on automated instru
SigNoz is a self-hosted observability platform that collects metrics via OpenTelemetry, stores them in a columnar database, and provides real-time dashboards and alerting, making it a strong match for server performance monitoring.
wgcloud is a comprehensive suite of monitoring and management tools designed for Linux servers, network devices, containers, and middleware. It functions as a centralized dashboard for tracking real-time hardware metrics, auditing the health of Docker and Kubernetes environments, and maintaining an IT asset management system for physical and cloud infrastructure. The platform is distinguished by its integrated remote administration capabilities, featuring a web-based SSH client for executing bulk commands and managing servers directly from a browser. It further differentiates itself with AI-d
wgCloud is a self-hostable, agent-based monitoring platform that provides real-time dashboards, multi-channel alerting, and multi-server support for Linux servers, containers, and network devices — directly covering the core requirements for server performance monitoring.
HertzBeat is an agentless monitoring platform designed to collect performance metrics from network devices, databases, and servers without requiring client software. It functions as an infrastructure monitoring dashboard, an alert management system, and a centralized log aggregator using the OpenTelemetry Protocol. The system utilizes a cloud-edge collection hierarchy to scale data gathering across clusters and isolated networks. It distinguishes itself with a flexible extensibility model, allowing users to define new monitoring workflows through configuration-based metric templates and custo
HertzBeat is a self-hosted monitoring platform that collects performance metrics from servers and infrastructure, offering real-time dashboards, alerting, and extensible metric templates—fitting your need for a performance monitoring tool, though its agentless design differs from the agent-based collection you requested.
Nezha is a multi-server infrastructure monitor and website uptime monitor that provides a centralized dashboard for tracking real-time resource utilization and system health. It functions as a protocol server and alerting engine, utilizing remote agents to collect telemetry data across multiple operating systems. The system distinguishes itself with a web-based remote administration interface, allowing users to execute maintenance commands and manage scheduled tasks on remote hosts via a browser-based terminal. It also integrates a Model Context Protocol server to provide a secure HTTP entry
Nezha is a self-hosted multi-server infrastructure monitor with remote agents for collecting real-time system metrics, a centralized dashboard, and built-in alerting — exactly the kind of agent-based, self-hostable performance monitoring tool you are looking for, with support for alerting and extensible agent provisioning.
Nightingale is a Prometheus-compatible monitoring and alerting platform designed to centralize telemetry management across multiple time-series databases. It functions as a multi-source alerting engine and metric data pipeline that ingests telemetry via remote write protocols and triggers alarms based on data from sources such as Prometheus, Elasticsearch, Loki, and ClickHouse. The system is distinguished by its automated alert healing system, which executes predefined scripts and RPC-based corrective actions when monitoring thresholds are breached. It supports distributed alert processing, a
Nightingale is a Prometheus-compatible monitoring and alerting platform that centralizes telemetry from multiple sources and provides real-time dashboards and alerting, which fits your need for a self-hosted server monitoring tool, though it relies on external agents for metric collection rather than bundling its own.
Glances is a cross-platform system monitoring tool designed to track real-time resource usage and hardware health metrics across diverse computing environments. It functions as a command-line utility that provides a unified view of system performance, identifying bottlenecks and maintaining infrastructure stability through a consistent abstraction layer that translates kernel calls into actionable data. The project distinguishes itself through its distributed capabilities, offering a web-based interface that enables remote access to live performance metrics from any device without requiring d
Glances is a self-hostable, cross-platform system monitor that collects real-time metrics via a local agent, offers a web dashboard for visualization, and supports alert triggers and multi-server remote access, fitting the core need for a server performance monitoring tool—though it lacks built-in time-series storage and its alerting is simpler than dedicated platforms.
Main repository for munin master / node / plugins
Munin is a classic open-source server monitoring system with an agent-based architecture, time-series RRD storage, built-in alerting, and extensive plugin support, making it a direct fit for self-hosted metric collection and visualization.
Linux-dash is a web-based system monitoring dashboard for Linux environments. It provides a visual interface for tracking hardware performance, system load, and real-time resource utilization. The project includes dedicated monitors for tracking the performance and resource usage of virtualized containers alongside a process manager for analyzing active system processes across the operating system. The dashboard covers several observability areas, including hardware performance monitoring for CPU and RAM, storage metrics for disk and swap space, and high-level system status overviews encompa
Linux-dash is a web-based dashboard for real-time Linux system monitoring, but it lacks agent-based multi-server support, alerting, and time-series storage, so it fits the category as a lightweight self-hosted option rather than a full-featured monitoring platform.
SkyWalking is a comprehensive observability stack and application performance monitoring platform. It functions as a distributed tracing system and an AI application monitor, providing a centralized suite for collecting and analyzing logs, metrics, and traces to maintain the health of containerized architectures. The platform distinguishes itself through a service topology visualizer that renders interactive maps of infrastructure dependencies and communication patterns. It also includes specialized capabilities for generative AI workflow observation to track the execution flow and performanc
SkyWalking is an observability platform that collects metrics via agents, provides real-time dashboards, and supports alerting—fitting your server monitoring needs, though its primary focus extends to distributed tracing and AI observability.
HyperDX is an OpenTelemetry observability platform that provides centralized log management, distributed tracing, and a self-hosted monitoring stack. It functions as a unified system for collecting, indexing, and visualizing logs, metrics, and traces from cloud and container environments. The platform distinguishes itself with specialized tooling for large language model monitoring and session replay, allowing user interactions in the browser to be linked to backend telemetry. It employs schema-less JSON parsing to index structured logs dynamically and uses source maps to resolve minified sta
HyperDX is an OpenTelemetry-based observability platform that covers the full stack of metrics, logs, and traces—so it handles agent-based metric collection, real-time dashboards, alerting, and time-series storage out of the box via its ClickHouse backend, all while being fully self-hostable and extensible through OpenTelemetry collectors, making it a capable tool for server performance monitoring.