Alertmanager
The deduplication, grouping, and routing layer for Prometheus alerts. Sends to PagerDuty, Slack, email, webhooks.
The deduplication, grouping, and routing layer for Prometheus alerts. Sends to PagerDuty, Slack, email, webhooks.
Apache APM platform for service topology, distributed tracing, metrics, logs, and cloud-native diagnostics.
Nagios-descended infrastructure monitoring with auto-discovery, pre-built service checks, and a polished web UI. Raw edition is GPL-2.0.
eBPF-based APM for Kubernetes. Zero-instrumentation service maps and SLO dashboards.
CNCF graduated lightweight log/metric forwarder. The default in-cluster log shipper for Kubernetes.
YAML-configured uptime monitoring with rich assertion DSL. Git-ops friendly.
Sentry-API-compatible error tracking, but MIT-licensed and lighter. Drop-in for Sentry SDKs.
Fast terminal and HTML dashboard for analyzing web server access logs in real time.
The de-facto open-source dashboard layer. Plug into Prometheus, Loki, Tempo, InfluxDB — one pane of glass.
Log aggregation inspired by Prometheus: index only labels, not full text. Cheap at scale.
Grafana Labs' horizontally scalable, multi-tenant Prometheus-compatible long-term metrics store.
Open-source on-call and incident-response — schedules, escalation policies, ChatOps. The closest OSS analog to PagerDuty.
Continuous profiling (CPU, memory, goroutines). Now part of the Grafana stack.
High-scale distributed tracing backend. Stores traces in cheap object storage (S3, GCS).
Mature log management and search with streams, pipelines, dashboards, and alerting. A practical Splunk alternative for log-heavy teams.
Cron-job and heartbeat monitoring for scheduled tasks. It alerts when an expected ping does not arrive.
Open-source Datadog alternative that correlates logs, metrics, traces, and session replays in one UI.
The CNCF graduated distributed tracing system, originally from Uber. The default open-source choice for traces.
Drop-in dashboard for Alertmanager — multi-cluster aggregation, richer grouping UI, on-call status, silence management.
Real-time per-second metrics with instant auto-discovery. Dashboards out-of-box, zero config.
Rust-based, columnar, object-storage-backed. Claims 140x lower storage cost than Elasticsearch for logs.
Apache-2.0 fork of Kibana, paired with OpenSearch. The de-facto self-hosted alternative to Elastic for logs visualization.
Vendor-neutral telemetry collector for receiving, processing, and exporting traces, metrics, and logs.
Rust-based log lake that writes Parquet to S3-compatible object storage. Compute-storage separated, indexless by design.
CNCF dashboarding tool with dashboards-as-code, static validation, and Kubernetes-native workflows. A GitOps-friendly Grafana alternative.
The CNCF-graduated, pull-based metrics store. Industry default for self-hosted metrics.
Rust search engine for logs and traces with sub-second queries directly against S3-compatible object storage.
Kubernetes-native observability that turns Prometheus alerts into actionable, context-enriched events with auto-remediation playbooks.
The upstream product itself — Sentry offers an official self-hosted docker-compose install under FSL license.
Full-stack, OpenTelemetry-native APM + logs + metrics. The closest single-tool Datadog alternative.
CNCF-graduated long-term storage and global query layer for Prometheus, backed by object storage.
Self-hosted uptime monitor + status page. The UptimeRobot replacement most engineers land on.
OpenTelemetry APM backed by ClickHouse. Lean alternative to SigNoz.
High-performance observability data pipeline. Sources, transforms, sinks — written in Rust by Datadog (open source).
Prometheus-compatible TSDB optimized for ingestion rate and storage cost. Single-binary by default, cluster mode for scale.
Classic agent-based monitoring suite for hosts, networks, and services. Pre-Prometheus old guard, still actively developed.
The original open-source distributed tracing system, born at Twitter in 2012. Single-binary simplicity, still maintained.
0 tools selected
No tools shortlisted.