Datadog
8-12 weeksA phased path for moving metrics, logs, traces, dashboards, and alerting away from Datadog without losing operational coverage.
Self-hosted observability, keyed on the SaaS you're replacing.
JavaScript disabled? Browse by category or SaaS you're replacing.
Engineering teams are leaving observability SaaS in record numbers. Datadog bills shocked the industry; New Relic’s 2023 pricing change sent waves of teams searching for alternatives.
This directory curates self-hosted, open-source observability tools — organized by the SaaS vendor they replace. No listicles. No affiliate-laundered opinions. Just structured entries with license, repo, deployment model, and honest notes.
Metrics · Logs · Traces · APM · Errors · Uptime · Profiling · All-in-One
A phased path for moving metrics, logs, traces, dashboards, and alerting away from Datadog without losing operational coverage.
A cautious path for replacing Honeycomb-style event analysis with OpenTelemetry pipelines, trace backends, and queryable logs.
A migration plan for teams replacing New Relic APM, logs, infrastructure metrics, and dashboards with OpenTelemetry-native tooling.
A focused plan for replacing hosted Sentry with self-hosted error tracking while preserving release, source map, alerting, and ownership workflows.
A practical plan for moving expensive Splunk log search and operational dashboards to open-source log pipelines and search backends.
The deduplication, grouping, and routing layer for Prometheus alerts. Sends to PagerDuty, Slack, email, webhooks.
Apache APM platform for service topology, distributed tracing, metrics, logs, and cloud-native diagnostics.
Nagios-descended infrastructure monitoring with auto-discovery, pre-built service checks, and a polished web UI. Raw edition is GPL-2.0.
eBPF-based APM for Kubernetes. Zero-instrumentation service maps and SLO dashboards.
CNCF graduated lightweight log/metric forwarder. The default in-cluster log shipper for Kubernetes.
YAML-configured uptime monitoring with rich assertion DSL. Git-ops friendly.
Sentry-API-compatible error tracking, but MIT-licensed and lighter. Drop-in for Sentry SDKs.
Fast terminal and HTML dashboard for analyzing web server access logs in real time.
The de-facto open-source dashboard layer. Plug into Prometheus, Loki, Tempo, InfluxDB — one pane of glass.
Log aggregation inspired by Prometheus: index only labels, not full text. Cheap at scale.
Grafana Labs' horizontally scalable, multi-tenant Prometheus-compatible long-term metrics store.
Open-source on-call and incident-response — schedules, escalation policies, ChatOps. The closest OSS analog to PagerDuty.
Continuous profiling (CPU, memory, goroutines). Now part of the Grafana stack.
High-scale distributed tracing backend. Stores traces in cheap object storage (S3, GCS).
Mature log management and search with streams, pipelines, dashboards, and alerting. A practical Splunk alternative for log-heavy teams.
Cron-job and heartbeat monitoring for scheduled tasks. It alerts when an expected ping does not arrive.
Open-source Datadog alternative that correlates logs, metrics, traces, and session replays in one UI.
The CNCF graduated distributed tracing system, originally from Uber. The default open-source choice for traces.
Drop-in dashboard for Alertmanager — multi-cluster aggregation, richer grouping UI, on-call status, silence management.
Real-time per-second metrics with instant auto-discovery. Dashboards out-of-box, zero config.
Rust-based, columnar, object-storage-backed. Claims 140x lower storage cost than Elasticsearch for logs.
Apache-2.0 fork of Kibana, paired with OpenSearch. The de-facto self-hosted alternative to Elastic for logs visualization.
Vendor-neutral telemetry collector for receiving, processing, and exporting traces, metrics, and logs.
Rust-based log lake that writes Parquet to S3-compatible object storage. Compute-storage separated, indexless by design.
CNCF dashboarding tool with dashboards-as-code, static validation, and Kubernetes-native workflows. A GitOps-friendly Grafana alternative.
The CNCF-graduated, pull-based metrics store. Industry default for self-hosted metrics.
Rust search engine for logs and traces with sub-second queries directly against S3-compatible object storage.
Kubernetes-native observability that turns Prometheus alerts into actionable, context-enriched events with auto-remediation playbooks.
The upstream product itself — Sentry offers an official self-hosted docker-compose install under FSL license.
Full-stack, OpenTelemetry-native APM + logs + metrics. The closest single-tool Datadog alternative.
CNCF-graduated long-term storage and global query layer for Prometheus, backed by object storage.
Self-hosted uptime monitor + status page. The UptimeRobot replacement most engineers land on.
OpenTelemetry APM backed by ClickHouse. Lean alternative to SigNoz.
High-performance observability data pipeline. Sources, transforms, sinks — written in Rust by Datadog (open source).
Prometheus-compatible TSDB optimized for ingestion rate and storage cost. Single-binary by default, cluster mode for scale.
Classic agent-based monitoring suite for hosts, networks, and services. Pre-Prometheus old guard, still actively developed.
The original open-source distributed tracing system, born at Twitter in 2012. Single-binary simplicity, still maintained.
0 tools selected
No tools shortlisted.