Open Source Monitoring Tools
Discover 52 open source Monitoring repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. Monitoring projects here are most often combined with Self Hosted, Observability and Docker. Last updated October 4, 2026.
52 repositories · updated October 4, 2026

APIPark: Manage AI Models and APIs Through One Gateway
APIPark is an open-source AI gateway and API developer portal for standardizing access to AI models and REST APIs. It helps teams manage API publishing, subscriptions, keys, usage monitoring, and model integrations in one place.

opik: Trace, Evaluate, and Monitor LLM Applications
Opik is a platform for tracing and evaluating LLM applications, RAG systems, and AI agents. Teams can use it to inspect workflows, run evaluations, and monitor deployments, either self-hosted or through Comet Cloud.

prometheus/prometheus: Monitor Systems with Time-Series Metrics
Prometheus collects and stores metrics from configured targets, lets teams query them with PromQL, and evaluates rules to surface issues or trigger alerts. It suits operators who want an autonomous monitoring system with flexible service discovery and metric labeling.

grafana: Explore and Visualize Observability Data
Grafana is a platform for querying, visualizing, and alerting on metrics, logs, and traces across data sources. It suits teams that need shared dashboards and a common way to explore operational data.

atlas: Discover and Visualize Network Infrastructure
Atlas scans Docker containers and nearby network hosts, then presents discovered devices in an interactive dashboard. It suits homelabs and small infrastructure setups that need scheduled discovery and visibility across multiple subnets.

ARIES: Automate Infrastructure Operations with AI
ARIES is a Python-based operations platform that combines LLM agents, infrastructure monitoring, and automated remediation. It targets teams managing servers, networks, and MQTT devices, with REST API access and webhook alerts.

nodemon: Automatically Restart Node.js Apps on File Changes
nodemon watches application files and restarts the running process when changes are detected. It helps Node.js developers shorten the edit-and-test loop without changing their application code.

opossum: Protect Node.js Calls with Circuit Breakers
Opossum is a JavaScript circuit breaker for Node.js asynchronous operations. It tracks failures and timeouts, fails fast when a dependency is unhealthy, and can run fallback logic while the service recovers.

handit.ai: Monitor and Automatically Fix AI Applications
Handit.ai monitors AI applications for quality and reliability issues, evaluates interactions, and can generate tested fixes as GitHub pull requests. It is aimed at teams operating production AI agents that want ongoing monitoring and a reviewable path from detection to code changes.

kafka-ui: Manage and Monitor Apache Kafka Clusters
A web UI for monitoring and managing Apache Kafka clusters from one place. It helps Kafka operators and developers inspect brokers, topics, consumer groups, and messages, and configure cluster workflows through a browser.

openllmetry: Observe LLM Applications with OpenTelemetry
OpenLLMetry adds OpenTelemetry-based tracing and metrics for LLM applications, providers, frameworks, and vector databases. It suits teams that want AI-specific observability in their existing telemetry stack rather than a separate monitoring format.

PatchMon: Manage and Patch Linux Server Fleets
PatchMon is a self-hosted platform for monitoring, securing, and patching server fleets from one interface. Its outbound-only agents support Linux, FreeBSD, and Windows, with patch approvals, compliance scans, inventory, and audit history.