SKILL.md packages that extend Claude Code, Cursor, Copilot, and other AI agents.
Tags

pup
Rust-based CLI that exposes Datadog APIs and a collection of agent skills for monitoring, logs, APM, security, and infrastructure management.

nightwatch
Practical guidance to configure Laravel Nightwatch: sampling, filtering, and redaction to balance observability, performance, and privacy.

agent-skills
Create, list, audit, and manage Datadog monitors with best-practice guidance to reduce alert noise and ensure actionable alerts.

linux-performance-annlysis-skill
Structured troubleshooting workflow to diagnose Linux performance bottlenecks across CPU, memory, I/O, and network with safety-first, low-intrusion defaults.

aidoc-flow-framework
Automate change-management (CHG) records end-to-end: detect changes, classify level, draft CHG, run audits/fixes, and prepare gate approval workflows.

aurora
Investigate CloudBees (Jenkins-compatible) builds, deployments, pipeline stages, logs, and test results to accelerate root cause analysis during incidents.

agents-in-a-box
SRE patterns for running and protecting autonomous agents: cost caps, circuit breakers, stall detection, observability, and runbooks to recover from incidents.

TerminalSkills Library
Configure PagerDuty services, escalation policies, on-call schedules and Event API alerts to automate incident workflows and routing.

skills
Run a short post-deploy monitor that captures screenshots, checks console errors, and compares performance against baselines to detect regressions and page fail

aura-frog
Generates and templates operational documentation: Architecture Decision Records (ADRs) and runbooks, with templates and principles for when and how to record d

neurofoo
Structured, blameless incident post-mortem template that produces an executive summary, timeline, root cause analysis, and actionable remediation items.

application-skills
Connect agents to PingBell via the Membrane CLI to manage calls, notifications, and incident response workflows for DevOps and SRE teams.

buiphucminhtam
Production-grade reliability framework covering SLOs, monitoring, chaos engineering, and incident management.

skills
Deploy and manage Grafana Mimir for scalable, long-term Prometheus and OpenTelemetry metrics storage with multi-tenancy support.

System Design Skills
Guidance on designing production monitoring, including metrics, logs, traces, SLOs, and health checks to prevent invisible failures.

grafana-lens
Full-spectrum Grafana observability: query metrics, logs, and traces, manage dashboards, and deploy Alloy data pipelines.

commonly-used-high-value-skills
Generates production-grade operational runbooks by analyzing codebases.