Best Open Source observability Libraries
A curated list of the most popular GitHub repositories tagged with observability. Select any project to visualize its architecture and dive into the codebase using RepoMind's AI engine.
#1netdata/netdata
The fastest path to AI-powered full stack observability, even for lean teams.
#2langfuse/langfuse
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23
#3SigNoz/signoz
SigNoz is an open-source observability platform native to OpenTelemetry with logs, traces and metrics in a single application. An open-source alternative to DataDog, NewRelic, etc. 🔥 🖥. 👉 Open source Application Performance Monitoring (APM) & Observability tool
#4cilium/cilium
eBPF-based Networking, Security, and Observability
#5mlflow/mlflow
The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models and data.
#6apache/skywalking
APM, Application Performance Monitoring System
#7elastic/kibana
Your window into all of your data
#8mikeroyal/Self-Hosting-Guide
Self-Hosting Guide. Learn all about locally hosting (on premises & private web servers) and managing software applications by yourself or your organization. Including Cloud, LLMs, WireGuard, Automation, Home Assistant, and Networking.
#9openobserve/openobserve
OpenObserve is an open-source observability platform for logs, metrics, traces, and frontend monitoring. A cost-effective alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage costs and single binary deployment.
#10openzipkin/zipkin
Zipkin is a distributed tracing system
#11kubesphere/kubesphere
The container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️
#12VictoriaMetrics/VictoriaMetrics
VictoriaMetrics: fast, cost-effective monitoring solution and time series database
#13Tracer-Cloud/opensre
Build your own AI SRE agents. The open source toolkit for the AI era.
#14upgundecha/howtheysre
A curated collection of publicly available resources on how technology and tech-savvy organizations around the world practice Site Reliability Engineering (SRE)
#15hyperdxio/hyperdx
Resolve production issues, fast. An open source observability platform unifying session replays, logs, metrics, traces and errors powered by ClickHouse and OpenTelemetry.
#16highlight/highlight
highlight.io: The open source, full-stack monitoring platform. Error monitoring, session replay, logging, distributed tracing, and more.
#17GreptimeTeam/greptimedb
The open-source observability database. One columnar engine for metrics, logs, and traces, on object storage.
#18grafana/mimir
Grafana Mimir provides horizontally scalable, highly available, multi-tenant, long-term storage for Prometheus.
#19micrometer-metrics/micrometer
An application observability facade for the most popular observability tools. Think SLF4J, but for observability.
#20parca-dev/parca
Continuous profiling for analysis of CPU and memory usage, down to the line number and throughout time. Saving infrastructure cost, improving performance, and increasing reliability.
#21rajnandan1/kener
Stunning status pages, batteries included!
#22DataDog/datadog-agent
Main repository for Datadog Agent
#23seakee/CPA-Manager-Plus
A self-hosted CPA / CLIProxyAPI management panel and AI gateway observability dashboard for requests, usage, cost, quota, failures, and account health.
#24perses/perses
The CNCF sandbox for observability visualisation. Already supports Prometheus, Tempo, Loki and Pyroscope - more data sources to come!
#25future-agi/future-agi
Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.
#26monoscope-tech/monoscope
Monoscope lets you ingest and explore your logs, traces and metrics. We store these in S3 compatible buckets. Query in natural language via LLMs.
#27percona/pmm
Percona Monitoring and Management: an open source database monitoring, observability and management tool
#28alibaba/anolisa
ANOLISA (Agentic Nexus Operating Layer & Interface System Architecture) | Agentic OS with runtime, security, observability, and Tokenless response compression for lower token usage and cost.
#29kuvasz-uptime/kuvasz
Kuvasz (pronounce as [ˈkuvɒs]) is an open-source uptime and SSL monitoring service, with multiple notification channels, status pages, IAC support via YAML, Prometheus integration, a complete REST API and many more!
#30nginx/agent
NGINX Agent provides an administrative entry point to remotely manage, configure and collect metrics and events from NGINX instances
#31laminlabs/lamindb
Open-source data management for multimodal AI. Query, trace, and govern with a lineage-native, format-agnostic lakehouse for agents and teams. With support for biological formats and registries by the creators of Scanpy. 🍊YC S22
#32Linuxfabrik/monitoring-plugins
Monitoring plugins for Icinga, Nagios & friends. Python 3.9+, all platforms. Smart defaults, auto-discovery, consistent cross-platform metrics, minimal dependencies.
#33Dynatrace/dynatrace-operator
Automate Kubernetes observability with Dynatrace
#34however-yir/knowledgeops-agent
Production-oriented Spring AI platform prototype for RAG, tool calling, async ingestion, JWT/RBAC, and observability.