ongridio/ongrid

An ops AI Agent that understands your infrastructure, finds the root cause, and fixes it — right from Slack, Telegram, Lark or DingTalk.

933Stars on the repository
34Mods indexed here, across every type
4d agoLast push, which is what freshness is scored on
AGPL-3.0Licence, which decides whether bodies are shown

ongridio/ongrid

Agent

An incident-investigation worker that follows a chain of causes back to the original source of a system alert or failure.

933 4d ago A 35 tokens AGPL-3.0

reporter

02

ongridio/ongrid

Agent

An operations-report writing worker that turns a ReportFacts JSON file into a structured ContentJSON report. ReportFacts contains system-calculated facts and numbers.

933 4d ago A 48 tokens AGPL-3.0

reviewer

03

ongridio/ongrid

Agent

A review worker for statically checking proposals that change or may destroy data.

933 4d ago A 21 tokens AGPL-3.0

specialist-compute

04

ongridio/ongrid

Agent

A computing and Linux-system troubleshooting assistant for CPU, memory, processes, and kernel behavior. It covers load, scheduling, context switches, out-of-memory events, NUMA, and kernel settings.

933 4d ago A 36 tokens AGPL-3.0

specialist-disk

05

ongridio/ongrid

Agent

A file-system and disk-capacity assistant for examining storage use on a computer. It covers tools and concepts such as du, find, stat, inodes, mounts, and large files.

933 4d ago A 29 tokens AGPL-3.0

specialist-network

06

ongridio/ongrid

Agent

A network troubleshooting assistant for Linux networking and related system tools. It covers OVS, netfilter, network namespaces, conntrack, bpftool, IP routing, firewalls, and network interfaces.

933 4d ago A 35 tokens AGPL-3.0

specialist-ops

07

ongridio/ongrid

Agent

An operations assistant for running and maintaining computer services. It covers service status, starting and restarting services, deployments, configuration, capacity, and scheduled tasks.

933 4d ago A 33 tokens AGPL-3.0

specialist-sre

08

ongridio/ongrid

Agent

An SRE and observability specialist for operating software systems. SRE means site reliability engineering, and observability means using system signals to understand health and failures.

933 4d ago A 35 tokens AGPL-3.0