A curated list of amazingly awesome tools, services and other shiny things for monitoring and analyze everything.
Complex infrastructure software
- Zabbix - Real-time monitoring of millions of metrics collected from tens of thousands of servers, virtual machines and network devices
- Nagios - Computer system, network and infrastructure monitoring software application.
- check_mk - Collection of extensions for Nagios.
- Opsview - Based on Nagios 4, Opsview Core is ideal for small IT and test environments.
- Centreon - IT infrastructure and application monitoring for service performance.
- Naemon - Network monitoring tool based on the Nagios 4 core with performance enhancements and new features.
- Icinga 2 - A Nagios like monitoring system, rewritten and expanded.
- openITCOCKPIT - Powerful open-source monitoring tool built upon Naemon or Nagios, featuring seamless integration with Grafana, an array of comprehensive reports, and visualizations.
- Sematext Cloud - Infrastructure and log monitoring with service and log auto-discovery. Basic plan is free.
- Middleware - Full-stack observability platform with infrastructure monitoring, Kubernetes monitoring, APM, logs, distributed tracing, real user monitoring, synthetic monitoring, and AI-assisted incident investigation.
- Fivenines - Server, uptime, cron and network monitoring. One platform. 5 first monitors are free.
Dashboards
- Grafana - The first really good dashboard for displaying metrics.
- Dash - A low-overhead monitoring web dashboard for a GNU/Linux machine.
- Munin - Networked resource monitoring tool.
- Adagios - Web based Nagios configuration interface.
- Thruk - Multibackend monitoring web interface with support for Naemon, Nagios, and Icinga.
- Uchiwa - Simple dashboard for the Sensu monitoring framework.
- Monit - Small Open Source utility for managing and monitoring Unix systems.
- Netdata - Troubleshoot slowdowns and anomalies in your infrastructure with thousands of metrics, interactive visualizations, and insightful health alarms.
- HomeLab Monitor - Self-hosted homelab dashboard in a single Docker container - per-container GPU/VRAM attribution, Docker health, systemd service status, and host vitals across multiple machines over SSH.
- KubeStellar Console - Multi-cluster Kubernetes monitoring dashboard with AI-powered operations and real-time observability across edge and cloud clusters.
- Kula - Lightweight, self-contained Linux server monitoring tool
Uptime and Synthetic Monitoring
External checks that request your endpoints from outside your own infrastructure
- Better Stack - Uptime monitoring bundled with log management and incident response, with on-call scheduling.
- BlueWave Uptime - Open-source, self-hosted monitoring tool built with React.js, Node.js, and MongoDB, designed to track server uptime, response times, and incidents in real-time with beautiful visualizations.
- Freshping - Free for 50 monitors, checked every 1 minutes, supports websocket monitoring
- Monitive - Free for 1 service, checked every 10 minutes with unlimited email & twitter alerts
- Oack - HTTP monitoring with TCP kernel telemetry, 6-phase latency breakdown, Server-Timing header capture, Cloudflare CDN enrichment, and built-in incident management with on-call scheduling.
- Checkly - Code-first synthetic monitoring for modern DevOps. Monitor your APIs and apps at a fraction of the price of legacy providers. Powered by a Monitoring as Code workflow and Playwright.
- Cronitor - Cron job and heartbeat monitoring alongside uptime checks, with schedule-aware alerting.
- Pingdom - Synthetic and real user monitoring with transaction checks from 100+ probe locations.
- Pulsetic - Uptime monitoring and status pages, checks run from 15 global locations.
- Safeship - Endpoint monitoring that validates the JSON body against a schema rather than only the status code, billed per check performed. Configurable from an AI editor over MCP.
- StatusCake - Uptime, page speed, server and SSL monitoring with Lighthouse data on standard plans.
- Uptime.com - 30+ check types including transaction monitoring and private location probes, with SLA reporting.
- UptimeRobot - Free for 50 monitors, checked every 5 minutes
- UpTime.onl - Free for 10 URLs, checked every 5 minutes
- UpTime360 - checked every 5 minutes. Monitor server, website, blacklist, custom services and publish status pages Get notified instantly on popular notification channels like - Slack, Twitter, Email, SMS (Twillo) and Pushover
- Kapient - Website monitoring for small businesses covering uptime, SSL, DNS, email deliverability, security, SEO, ADA compliance, and Google Business Profile, with AI-generated fix instructions tailored to your tech stack.
- elmah.io - Uptime monitoring combined with application error logging
- StatusList.app - Uptime monitoring with debug details and hosted status page in one dashboard
- ePulz.io - EU-hosted, GDPR-friendly uptime monitoring with public status pages, SSL and domain expiry, heartbeats, visual checks and alerts via email/Telegram/Slack/MS Teams. 14 languages.
- Sematext Synthetics - Website uptime, API, and SSL certificate monitoring. Includes status pages and scriptable multi-page user transaction monitoring, etc.
- Tianji - All-in-One Insight Hub
- SSL Certificate Monitor - Open-source SSL/TLS certificate expiry monitoring tool with web UI and REST API. Checks certificate validity and days until expiration.
- Uptime Kuma - An easy-to-use self-hosted monitoring tool.
- Phare - Free 100k monitoring events per months, 30s intervals, unlimited users, incident management, and sleek status pages.
- API Status Check - Free real-time status monitoring dashboard for 114+ developer APIs including AWS, Stripe, GitHub, and OpenAI.
- Oh Dear - Monitoring for uptime, performance, SSL certificates, broken links, and DNS, with hosted status pages
- DevHelm - Developer-first uptime monitoring (HTTP, DNS, TCP, ICMP, heartbeat) with dependency intelligence for 80+ providers, hosted status pages, and config-as-code. Free tier: 50 monitors.
- Uptimeify.io - Reliable and simple website and API monitoring with instant alerts and status pages.
- VisualSentinel - Website monitoring with uptime, visual change detection (screenshot diff of the live site), SSL, DNS, performance, and content checks. Free tier for 3 monitors, checks from EU and US. Alerts via Slack, Discord, WhatsApp, Telegram, PagerDuty and webhooks.
- Crontiq - Free cron job monitoring with automatic JSON metric extraction and zero-config anomaly detection. 20 monitors free forever.
- Spork - Uptime monitoring with built-in status pages, CLI, and Terraform provider. Multi-region consensus checks from $4/mo.
- FlareWarden — Uptime, content, dependency, and SSL monitoring with multi-region verification and status pages. Free plan includes 15 monitors, 5-minute checks, and 90 days of history.
- Hyperping - Uptime, API, cron, and server monitoring from 18 locations, with Playwright browser checks, on-call scheduling, and hosted status pages.
- UpWatchr - Free native Windows desktop app for uptime monitoring of websites and services. Local-first: checks run from your own machine, no cloud and no account.
- sunwatch - Crypto-paid uptime monitoring for side projects. Pay per monitor with USDC on Base; webhook alerts on down/up state changes.
- Drumbeats - Cron, heartbeat, and uptime monitoring with incident management and status pages. Free for up to 50 monitors, 200K Beats/mo, 1-minute checks, no credit card. One curl ping to instrument a job, no agent or SDK.
- Superhighway - Web API whose
/scrapeendpoint turns any URL into clean Markdown, useful for building your own content/page-change monitoring (e.g. competitor pricing, changelogs, pages without an RSS feed). See the change-detection tutorial: scrape, SHA-256 hash comparison, LLM-summarize what changed, and schedule with cron or GitHub Actions. Free API key or pay-per-call.
Application Performance monitoring
- NewRelic - Complex service for both application and infrastructure monitoring
- Last9 - OpenTelemetry-native observability platform for APM, metrics, logs, and traces, built to handle high-cardinality data at scale.
- Uptrace - OpenTelemetry-native APM.
- DataDog - Complex service for both application and infrastructure monitoring
- OverOps - OverOps provides Automated Root Cause (ARC) analysis to reduce the time to identify and fix critical production application errors.
- AppSignal - Catch errors, track performance, monitor hosts, detect anomalies — all in one tool.
- BitDive - APM for Java/Kotlin with distributed tracing, method-level profiling, and performance metrics.
- Middleware - Middleware's APM helps you troubleshoot issues in real-time, optimize performance, improve user experience, and reduce downtime.
- PerfScope - Advanced Python performance profiler with decorator-based setup, call tree visualization, memory tracking, and multi-format report export.
- Muscula - Lightweight application monitoring platform for developers with AI-native error reporting and debugging.
- groundcover - eBPF-based observability platform for Kubernetes with logs, metrics, traces, and APM; deployed inside the user's own cloud (BYOC).
- CoreDash - Real user monitoring for Core Web Vitals (LCP, INP, CLS, TTFB, FCP) with element level attribution, request waterfalls, LoAF data, and a built in MCP server so AI agents can query live performance data. EU hosted, GDPR compliant.
- Matomo - Take back control with Matomo – a powerful web analytics platform that gives you 100% data ownership.
- Heap Analytics - Easy event tracking without coding
- Screpy - Screpy is a web analyzer and monitoring tool. Its powered by Google Lighthouse.
- PageGuard - Free all-in-one website health scanner powered by Lighthouse. Monitors performance, SEO, accessibility, and best practices with AI-generated action plans and a free REST API.
- Shynet - Modern, privacy-friendly, and cookie-free web analytics.
- Tianji - All-in-One Insight Hub
- API Status Check - Real-time status dashboard for 160+ third-party APIs including AI platforms (OpenAI, Anthropic), cloud providers (AWS, Vercel), payments (Stripe, PayPal), and developer tools (GitHub, Supabase). Includes embeddable status badges.
- DownStatus - Free JSON API providing real-time status for GitHub, AWS, Discord, Cloudflare, Stripe, and 90+ popular services.
- IncidentHub - Status page aggregator for monitoring SaaS and cloud providers.
- OpenChainBench - Open live benchmarks of blockchain infrastructure APIs and RPC providers (latency p50/p90/p99, success rate, multi-region), built on a public Prometheus; all data CC-BY-4.0.
- StatusGator - Cloud service monitoring, aggregating status pages of cloud services into a single dashboard.
- Voidly Check - GitHub Action that verifies your services aren't blocked in target countries (Iran, China, Russia, etc.) using real OONI measurements. Backed by
api.voidly.ai(19.6M samples). - ClaudeDown - Real-time Claude AI complaint tracker using Twitter/X sentiment data to detect outages before official status pages.
- Apitally - Analytics, request logging and monitoring for REST APIs with a focus on simplicity and data privacy.
- BurnRate - AI coding cost analytics CLI that tracks usage and spend across Claude Code, Cursor, Copilot, Windsurf, Aider, Cline, and Codex. Includes optimization rules and rate limit monitoring.
- ReqKey - API key authentication, usage credits, rate limiting, and request analytics for REST APIs via a single verify call. SDKs for Python, Node.js, Go, Rust, PHP, .NET, and Java, with a free tier of 100,000 requests per month.
- Honeybadger - Monitor application errors, performance, uptime, and logs in one simple tool for developers.
- Sentry - Application monitoring, event logging and aggregation.
- Bugsnag - Application monitoring, event logging and aggregation.
- Banshee - Real-time anomalies(outliers) detection system for periodic metrics
- Brubeck - Statsd-compatible stats aggregator written in C
- Loggly - Aggregate & analyze logs from any source
- Last9 - OpenTelemetry-native observability platform with log collection, search, and analysis.
- Logit.io - Centralise logs and metrics using the ELK Stack, Grafana & Open Distro.
- Sematext Logs - Log monitoring with log auto-discovery and alerting; comes with out of the box dashboards, pipelines for transforming, masking, dropping, sampling log events and more. Basic plan is free.
- Moira - Most powerful alerting system, backed by Graphite.
- Alerta - Distributed, scaleable and flexible monitoring system.
- Flapjack - Monitoring notification routing & event processing system.
- Seyren - An alerting dashboard for Graphite.
Tools for databases
- Anemometer - MySQL Slow Query Monitor
- QuerySharp - AI-Powered PostgreSQL performance monitoring. Slow query analysis, index recommendations, query rewriting.
Databases
- Graphite - More, than a time series database. And so awesome using with Grafana.
- Prometheus - Time series database for real-time monitoring and alerting
- InfluxDB - Time-series database built from the ground up to handle high write and query loads
- Levitate - A Managed Time Series Metrics and Events Warehouse built to handle High Cardinality data.
- Cacti - Web-based network monitoring and graphing tool.
- dish - A lightweight monitoring service that efficiently checks socket connections and can be configured remotely.
- NetHawk - Real-time network traffic analysis TUI built in Go. Features bandwidth monitoring, protocol breakdown, top talkers, and DDoS attack detection.
- Observium - SNMP monitoring for servers and networking devices. Runs on linux.
- Smokeping - SmokePing is a deluxe latency measurement tool.
- LibreNMS - Fork of Observium.
- Fluere - Versatile network interface monitoring and analysis tool, capable of capturing network packets in pcap format, NetFlow data. supports lua based plugins
- Flowtriq - DDoS detection for firewalls and routers (pfSense, OPNsense, VyOS) via NetFlow/sFlow export. Detects volumetric floods, SYN floods, amplification attacks, and provides real-time alerting with automated mitigation.
- Awesome Performance Engineering - Observability and performance testing tools for performance engineering.
