Best for: Knowing a service is down before anyone tells you.
The single highest-value self-hosted service in this entire directory. It handles HTTP, TCP, ping, DNS and push monitors, alerts through dozens of channels including Telegram and email, and publishes a public status page you can share. It is light enough to run on the same box it monitors, which is a slight limitation and also the reason people actually install it.
Developer Tools
MIT
Beginner
★ 91,997
Replaces: UptimeRobot, Pingdom, Better Stack
Full review of Uptime Kuma →
Best for: Configuration-as-code health checks with a status page.
The declarative alternative: checks defined in a YAML file, with conditions on status codes, response bodies and response time. For anyone who already keeps configuration in version control, this is a far better fit than a click-through interface, and it produces a clean status page as a side effect.
Monitoring & Analytics
Apache-2.0
Beginner
★ 12,217
Replaces: UptimeRobot, Pingdom, Statuspage
Full review of Gatus →
Best for: Instant, zero-configuration insight into a machine's current state.
Worth installing on every machine you own, because it requires no configuration and immediately shows per-second CPU, memory, disk, network and per-container detail. Its strength is instant diagnosis — you open it during a problem and the cause is usually visible within seconds. Long-term retention and alerting are the weaker parts.
Monitoring & Analytics
GPL-3.0
Beginner
★ 80,777
Replaces: Datadog, Nagios, Zabbix
Full review of Netdata →
Best for: Dashboards and visualisation over any metric source.
Where metrics become readable: one place to build dashboards over Prometheus, Loki, InfluxDB, PostgreSQL and more. It is not itself a data store, and that is precisely why it pairs with almost everything else on this list. For most people it is the single most useful long-term addition to a homelab.
Monitoring & Analytics
AGPL-3.0
Intermediate
★ 77,041
Replaces: Datadog, New Relic, Grafana Cloud
Full review of Grafana →
Best for: Collecting and querying metrics across many services.
The de facto standard for metrics collection, with a query language that has become the vocabulary of the whole discipline and an ecosystem of exporters for essentially everything. Adopting it is adopting a concept rather than a product. Budget for the storage, and configure retention deliberately.
Monitoring & Analytics
Apache-2.0
Advanced
★ 66,346
Replaces: Datadog, New Relic, AWS CloudWatch
Full review of Prometheus →
Best for: A quick per-machine overview from a terminal.
A terminal-based system monitor with a web mode, ideal for looking at one machine quickly without permanently running a metrics stack. Lower commitment than Netdata and easier to read than raw `top` output.
Monitoring & Analytics
LGPL-3.0
Beginner
★ 33,715
Replaces: htop, top, Netdata
Full review of Glances →
Best for: Centralised logs without the cost of full-text indexing.
Log aggregation that indexes labels rather than the full text of every line, which makes it dramatically cheaper to run than a full-text log index. The trade is that queries are slower unless you have chosen your labels well. It is the pragmatic answer to "I want logs in Grafana next to my metrics".
Monitoring & Analytics
AGPL-3.0
Advanced
★ 28,983
Replaces: Elasticsearch, Splunk, Datadog Logs
Full review of Grafana Loki →
Best for: Reading container logs quickly from a browser.
Live Docker container logs in a browser, with essentially no setup. It will not replace a log aggregation system, but it covers the most common day-to-day need — seeing what a misbehaving container is saying — better and faster than almost anything else.
Monitoring & Analytics
MIT
Beginner
★ 14,528
Replaces: Datadog Logs, Papertrail, Docker Desktop
Full review of Dozzle →
Best for: Traces, metrics and logs together for application debugging.
The strongest open-source observability platform here, with traces, metrics and logs in one interface built on OpenTelemetry. Adopt it when you have an actual distributed-systems problem, because instrumenting with OpenTelemetry from the start is what keeps you from being locked to a vendor later.
Developer Tools
Apache-2.0
Advanced
★ 32,253
Replaces: Datadog, New Relic, Grafana Cloud
Full review of SigNoz →
Best for: Understanding what is actually happening on your network.
Network-layer visibility that nothing else here provides: per-host and per-protocol traffic, flows, and history. Worth running temporarily when you need to answer "what on my network is using all the bandwidth", even if you do not keep it on permanently.
Monitoring & Analytics
GPL-3.0
Advanced
★ 8,220
Replaces: Darktrace, PRTG, SolarWinds NPM
Full review of ntopng →
Best for: Longer metric retention on limited hardware.
A drop-in replacement for Prometheus storage that is markedly more efficient in both disk space and memory for the same data, and accepts the same remote-write protocol. If your Prometheus instance is straining its VPS, this is the least disruptive fix available.
Monitoring & Analytics
Apache-2.0
Intermediate
★ 17,800
Replaces: Prometheus, InfluxDB, Thanos
Full review of VictoriaMetrics →
Best for: Sensor and IoT metrics with high write volume.
A time-series database with a gentler query language than PromQL and strong support for the IoT and sensor workloads that dominate home labs. The right choice when your metrics come from devices and scripts rather than from exporters.
Monitoring & Analytics
MIT
Intermediate
★ 31,759
Replaces: Prometheus, TimescaleDB, VictoriaMetrics
Full review of InfluxDB →
Best for: Routing, grouping and silencing alerts once you have many rules.
Prometheus detects conditions; Alertmanager decides who hears about them, how often, and how duplicates are collapsed. Without it, an incident produces two hundred identical notifications and you stop reading them. Essential once you go beyond a handful of alert rules.
Monitoring & Analytics
Apache-2.0
Intermediate
★ 8,636
Replaces: PagerDuty, Opsgenie, VictorOps
Full review of Alertmanager →