All tags
Catalogue tag

#sre

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

42 records
Tagged “sre”Ranked by health index
npm
96Exceptionalhealth index
chaterm/Chaterm
Open source AI terminal for cloud and infrastructure management, enabling you to deploy, troubleshoot, and automate services using natural language and intelligent agents.
TypeScript · Vue★ 3,036↓ 0/moJul 22, 2026
Custom licenseJul 22, 2026 · metrics 2.10.0
npm
96Exceptionalhealth index
jaegertracing/jaeger-ui
Web UI for Jaeger
TypeScript · JavaScript★ 1,513Jul 28, 2026
Apache-2.0Jul 28, 2026 · metrics 2.10.0
Maven · npm
94Exceptionalhealth index
rundeck/rundeck
Enable Self-Service Operations: Give specific users access to your existing tools, services, and scripts
Groovy · Java★ 6,279Aug 28, 2026
Apache-2.0Aug 28, 2026 · metrics 2.10.0
Go
91Excellenthealth index
kubeshark/kubeshark
eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.
Go★ 12.1KAug 27, 2026
Apache-2.0Aug 27, 2026 · metrics 2.10.0
Go · npm
90Excellenthealth index
coroot/coroot
Coroot is an open-source observability and APM tool with AI-powered Root Cause Analysis. It combines metrics, logs, traces, continuous profiling, and SLO-based alerting with predefined dashboards and inspections.
Go · Vue★ 7,858Aug 2, 2026
Apache-2.0Aug 2, 2026 · metrics 2.10.0
Go
89Excellenthealth index
jordigilh/kubernaut
An AIOps platform that closes the loop from Kubernetes alert to automated remediation: An AI Agent investigates via MCP tools, selects a fix from a pre-seeded workflow catalog and delegates execution (K8s Job, Tekton, Ansible) — or escalates with a full RCA. Approval gates, OPA policies, and audit trails keep humans in control.
Go★ 28Aug 29, 2026
Apache-2.0Aug 29, 2026 · metrics 2.10.0
Go · npm
89Excellenthealth index
rootlyhq/terraform-provider-rootly
Terraform provider for Rootly - manage incident management, on-call schedules, workflows, and alerts as code
Go★ 22Jul 17, 2026
MPL-2.0Jul 17, 2026 · metrics 2.10.0
Go · npm
87Excellenthealth index
VersusControl/versus-incident
Versus Incident is the self-hosted AI SRE agent. It learns what your system normally look like and escalates only what is new or unexpected issues — routing to your chat channels and on-call platform.
Go · TypeScript★ 645Jul 18, 2026
MITJul 18, 2026 · metrics 2.10.0
Go · npm
87Excellenthealth index
nobl9/sloctl
A command line tool to cast SLO spells 🪄
Go · Shell★ 39Aug 11, 2026
MPL-2.0Aug 11, 2026 · metrics 2.10.0
Go
84Excellenthealth index
platform-engineering-labs/formae
Infrastructure-as-Code Platform Built for the Future
Go★ 761Jul 20, 2026
Custom licenseJul 20, 2026 · metrics 2.10.0
npm
81Excellenthealth index
scitix/siclaw
AI-powered SRE platform — read-only infrastructure diagnostics with deep investigation, security governance, and team collaboration
TypeScript · Python★ 223↓ 45/moJul 28, 2026
Apache-2.0Jul 28, 2026 · metrics 2.10.0
Go · npm
80Excellenthealth index
FluidifyAI/Regen
Open-source oncall management Alerts, Incidents, AI post-mortems. Self-hosted alternative to PagerDuty & incident.io. BYO-AI, Works with Prometheus, Grafana, Datadog, Slack, and Teams
Go · TypeScript★ 74Aug 22, 2026
Custom licenseAug 22, 2026 · metrics 2.10.0
Go
80Excellenthealth index
nudgebee/node-agent
Per-node observability agent for Kubernetes and Linux hosts. Gathers container and host metrics, logs, and L7 traffic via eBPF; exports to Prometheus and OpenTelemetry. Includes LLM API observability.
Go★ 0Jul 22, 2026
Apache-2.0Jul 22, 2026 · metrics 2.10.0
npm
78Goodhealth index
ThoTischner/observability-mcp
Unified observability gateway for AI agents — one MCP server for Prometheus, Loki, and any backend, with cross-signal anomaly detection and a built-in Web UI.
TypeScript · JavaScript · HTML★ 6↓ 2,001/moJul 25, 2026
Apache-2.0Jul 25, 2026 · metrics 2.10.0
Go · npm
78Goodhealth index
ongridio/ongrid
An ops AI Agent that understands your infrastructure, finds the root cause, and fixes it — right from Slack, Telegram, Lark or DingTalk.
Go · TypeScript★ 492Jul 19, 2026
AGPL-3.0Jul 19, 2026 · metrics 2.10.0
npm
77Goodhealth index
swenyai/sweny
AI-powered engineering workflows — Learn from any source, Act through any tool, Report through any channel
TypeScript★ 3↓ 3,274/moAug 25, 2026
MITAug 25, 2026 · metrics 2.10.0
Go
75Goodhealth index
Smana/runlore
The self-improving SRE agent
Go★ 8Jul 15, 2026
Apache-2.0Jul 15, 2026 · metrics 2.10.0
Go
73Goodhealth index
fclairamb/solidping
Distributed, self-hostable uptime monitoring: 39 check types, multi-region workers, private agents, status pages, incidents and on-call escalation — in a single Go binary.
Go · TypeScript★ 2Aug 22, 2026
AGPL-3.0Aug 22, 2026 · metrics 2.10.0
Go
71Goodhealth index
revelara-ai/rvl-cli
Revelara CLI — connect your codebase to the Revelara reliability risk platform
Go★ 1Jul 28, 2026
Apache-2.0Jul 28, 2026 · metrics 2.10.0
npm
69Goodhealth index
JaniAnttonen/winston-loki
Grafana Loki transport for the nodejs logging library Winston.
JavaScript★ 170↓ 973.4K/moJul 18, 2026
MITJul 18, 2026 · metrics 2.10.0
Go
69Goodhealth index
conallob/o11y-analysis-tools
Various static analysis and testing tools for managing PromQL compatible monitoring stacks
Go★ 1Jul 27, 2026
BSD-3-ClauseJul 27, 2026 · metrics 2.10.0
Go
69Goodhealth index
hrodrig/groot
GROOT — Kubernetes cluster diagnostics CLI Collect nodes, events, pod logs, describes and more with one command. Fast, parallel, configurable, and outputs a clean .tar.gz archive. Perfect for incident response and troubleshooting.
Go★ 11Jul 17, 2026
MITJul 17, 2026 · metrics 2.10.0
Go · npm
67Goodhealth index
manishchaudhary101/kube-argus
Real-time Kubernetes dashboard for SREs — live cluster state, drain wizard, YAML editor, just-in-time exec access, cost analysis, and AI diagnosis in a single binary
TypeScript · Go★ 126Aug 22, 2026
Apache-2.0Aug 22, 2026 · metrics 2.10.0
Go
65Goodhealth index
hrodrig/kzero
Go CLI for declarative Kubernetes pipelines (down, up, reset). Turn start-over into a checked-in playbook: return workloads to a known initialization state—ordered scale-down and bring-up for Deployments, StatefulSets, Helm release steps, and custom scripts. Includes API watchdog, throttled progress logs, and delivery-visible notify events.
Go★ 3Jul 23, 2026
MITJul 23, 2026 · metrics 2.10.0
Go · npm
65Goodhealth index
imneeteeshyadav98/kubepreflight
Kubernetes upgrade-readiness CLI and Console for detecting deprecated APIs, webhook blockers, PDB risks, unhealthy workloads, node skew, and upgrade action plans.
Go · TypeScript★ 0Jul 26, 2026
Apache-2.0Jul 26, 2026 · metrics 2.10.0
Go
60Moderatehealth index
kordloom/switchtender
One Go binary that runs Ansible, Terraform, OpenTofu, Bash, PowerShell, Python, and Go across a fleet. Live host-by-task matrix, hash-chained audit you can verify offline, enforced approvals. No Kubernetes, no Postgres, no Redis. Migrate from AWX, Semaphore, Ascender, or Ansible Automation Platform in one command.
Go · HTML★ 0Aug 3, 2026
Custom licenseAug 3, 2026 · metrics 2.10.0
Go
59Moderatehealth index
kirilurbonas/FireDrill
Fire drills for your backups — restores real backups into disposable sandboxes, verifies the data, measures RTO/RPO, emits signed audit evidence (SOC 2 / ISO 27001)
Go★ 0Jul 26, 2026
Apache-2.0Jul 26, 2026 · metrics 2.10.0
Go · npm
59Moderatehealth index
projecthelena/warden
Open-source uptime monitoring built in Go. Multi-zone checks, status pages, unlimited team members — the production-grade upgrade from Uptime Kuma.
TypeScript · Go★ 3Aug 13, 2026
AGPL-3.0Aug 13, 2026 · metrics 2.10.0
59Moderatehealth index
tedilabs/terraform-aws-security
🌳 A sustainable Terraform Package which creates Security resources on AWS
HCL★ 19Aug 8, 2026
Apache-2.0Aug 8, 2026 · metrics 2.10.0