Home / Stacks / Incident Response & On-Call
🚨

Incident Response & On-Call

Detect, route, and resolve production incidents — Prometheus Alertmanager routing, PagerDuty on-call and escalation, Uptime Kuma availability monitoring, Prometheus metrics, Grafana dashboards, and an SLO-driven observability baseline

6 skills · Works with Claude Code, Codex, Cursor & more

⚙️ Engineering
RARE

PagerDuty

Set up incident management, on-call scheduling, and escalation with PagerDuty. Create services and escalation policies, wire integrations from monitoring tools, manage on-call rotations, and automate incident response workflows via the API.

Community 1.5K
Scanned
pagerduty incident-management on-call
mkdir -p ~/.claude/skills/pagerduty && curl -fsSL https://raw.githubusercontent.com/TerminalSkills/skills/main/skills/pagerduty/SKILL.md -o ~/.claude/skills/pagerduty/SKILL.md
⚙️ Engineering
RARE

Prometheus Alertmanager

Route, group, silence, and deliver alerts with Prometheus Alertmanager. Build routing trees, wire receivers for Slack, PagerDuty, and email, manage silences and inhibition rules, and debug flaky alert-delivery pipelines.

Community 2.6K
Scanned
alertmanager prometheus alerting
mkdir -p ~/.claude/skills/alertmanager && curl -fsSL https://raw.githubusercontent.com/TerminalSkills/skills/main/skills/prometheus-alertmanager/SKILL.md -o ~/.claude/skills/alertmanager/SKILL.md
⚙️ Engineering
RARE

Uptime Kuma

Self-hosted uptime monitoring with Uptime Kuma. Set up HTTP, TCP, DNS, Docker, and keyword checks; send multi-channel alerts to Slack, Telegram, Discord, and email; publish public status pages; and track SLA and response-time metrics.

Community 2.8K
Scanned
uptime-kuma monitoring status-page
mkdir -p ~/.claude/skills/uptime-kuma && curl -fsSL https://raw.githubusercontent.com/TerminalSkills/skills/main/skills/uptime-kuma/SKILL.md -o ~/.claude/skills/uptime-kuma/SKILL.md
⚙️ Engineering
RARE

Prometheus Monitoring

Query and operate Prometheus via its HTTP API. Writes PromQL for instant and range queries, inspects targets, alerts, and rules, accesses series and label metadata, and runs TSDB admin operations to troubleshoot monitoring infrastructure.

Community 2.6K
Scanned
prometheus promql monitoring
mkdir -p ~/.claude/skills/prometheus && curl -fsSL https://raw.githubusercontent.com/julianobarbosa/claude-code-skills/main/skills/prometheus/SKILL.md -o ~/.claude/skills/prometheus/SKILL.md
⚙️ Engineering
RARE

Grafana Dashboard Builder

Design and generate Grafana dashboards with PromQL and LogQL queries. Creates panels for system metrics, application KPIs, SLO tracking, and alerting rules with proper thresholds and visualization types.

Community 2.6K
Scanned
grafana monitoring dashboards
mkdir -p ~/.claude/skills/grafana-dashboards && curl -fsSL https://raw.githubusercontent.com/rampstackco/claude-skills/HEAD/skills/monitoring-and-alerting/SKILL.md -o ~/.claude/skills/grafana-dashboards/SKILL.md
⚙️ Engineering
RARE

Monitoring & Observability Strategy

Design end-to-end observability: metrics, logs, and traces. Applies the Four Golden Signals and RED/USE methods, stands up Prometheus/Grafana/Loki, instruments with OpenTelemetry, calculates SLOs and error budgets, and optimizes Datadog cost or migrates to an open-source stack.

Community 3.4K
Scanned
observability monitoring slo
mkdir -p ~/.claude/skills/monitoring-observability && curl -fsSL https://raw.githubusercontent.com/ahmedasmar/devops-claude-skills/main/monitoring-observability/skills/SKILL.md -o ~/.claude/skills/monitoring-observability/SKILL.md

More Stacks

View all stacks →

Added to wishlist