Reliability

Browse Reliability agent skills in Security and compare related workflows, tools, and use cases.

15 skills
C
uptimerobot-automation

by ComposioHQ

uptimerobot-automation helps AI agents run UptimeRobot Monitoring workflows through Composio Rube MCP. Use it to install the skill, connect Rube MCP, verify the UptimeRobot connection, discover current tool schemas with RUBE_SEARCH_TOOLS, and execute monitor tasks safely.

Monitoring
Favorites 0GitHub 67.5k
C
updown-io-automation

by ComposioHQ

updown-io-automation helps agents run Updown.io monitoring tasks through Composio Rube MCP. Learn setup requirements, RUBE_SEARCH_TOOLS discovery, connection checks, and safe usage patterns.

Monitoring
Favorites 0GitHub 67.5k
W
error-handling-patterns

by wshobson

error-handling-patterns helps teams choose exceptions vs Result types, classify failures, propagate context, and design graceful degradation for more reliable APIs and services.

Reliability
Favorites 1GitHub 32.6k
W
python-resilience

by wshobson

python-resilience is a guidance skill for safer Python failure handling with retries, exponential backoff, jitter, timeouts, and bounded retry windows. Use it to install practical resilience patterns for external calls and apply tenacity-style wrappers with clearer retry rules.

Reliability
Favorites 0GitHub 32.6k
W
slo-implementation

by wshobson

Use the slo-implementation skill to define SLIs, SLOs, error budgets, and burn-rate alerts for Reliability work. It helps teams turn service goals into measurable targets with PromQL-style examples and practical guidance from SKILL.md.

Reliability
Favorites 0GitHub 32.6k
W
istio-traffic-management

by wshobson

istio-traffic-management helps teams draft Istio traffic policies like VirtualService, DestinationRule, Gateway, and ServiceEntry for canary, retries, circuit breaking, and mirroring. Use it to translate deployment intent into clear routing and resilience manifests with practical prompts and review checks.

Deployment
Favorites 0GitHub 32.6k
W
linkerd-patterns

by wshobson

linkerd-patterns helps teams apply Linkerd patterns for Kubernetes workloads, including mTLS, sidecar injection, traffic splits, retries, timeouts, service profiles, and multi-cluster planning for Deployment-based rollouts.

Deployment
Favorites 0GitHub 32.6k
W
on-call-handoff-patterns

by wshobson

Learn the on-call-handoff-patterns skill for reliable shift transitions. Use it to structure incident handoffs, capture active issues, recent changes, escalation state, and next actions for Reliability teams.

Reliability
Favorites 0GitHub 32.5k
W
incident-runbook-templates

by wshobson

incident-runbook-templates helps teams create structured incident response runbooks with clear triage, mitigation, escalation, communication, and recovery steps for outages and operational Playbooks.

Playbooks
Favorites 0GitHub 32.5k
A
runbook-generator

by alirezarezvani

runbook-generator creates operational runbook drafts for services using a Python CLI and templates for deployment, health checks, rollback, incident response, maintenance, and validation. Useful for SRE, DevOps, and Technical Writing teams standardizing on-call procedures.

Technical Writing
Favorites 0GitHub 22.2k
A
observability-designer

by alirezarezvani

observability-designer helps SRE and platform teams design observability for APIs and services with dashboard generation, alert-noise analysis, and lightweight SLI/SLO scaffolds using included Python scripts, samples, and references.

Observability
Favorites 0GitHub 22.2k
A
feature-flags-architect

by alirezarezvani

feature-flags-architect helps teams plan, audit, and clean up feature flags for progressive delivery. Use it for rollout plans, kill-switch checks, stale flag scans, provider comparisons, and lifecycle guidance across LaunchDarkly, GrowthBook, Statsig, Unleash, Flipt, or DIY systems.

Deployment
Favorites 0GitHub 22.2k
A
feature-flags-architect

by alirezarezvani

feature-flags-architect is a Software Architecture skill for designing, rolling out, auditing, and retiring feature flags. It includes stdlib Python scripts for flag debt scanning, rollout planning, and kill-switch audits, plus references for flag taxonomy, lifecycle, rollout strategies, and provider trade-offs.

Software Architecture
Favorites 0GitHub 22.1k
A
vpe-advisor

by alirezarezvani

vpe-advisor is a VP of Engineering operating-advice skill for startup delivery throughput, hiring funnel health, team structure, and production discipline. Use its references and Python tools to analyze DORA metrics, hiring gaps, manager triggers, on-call practices, and strategic planning tradeoffs.

Strategic Planning
Favorites 0GitHub 22.1k
S
upgrade-stripe

by stripe

upgrade-stripe guide for upgrading Stripe API versions, server-side SDKs, Stripe.js, and mobile SDKs in real codebases, with practical steps for Backend Development.

Backend Development
Favorites 0GitHub 1.5k