Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
daemon-blockint-tech avatar

Sla Slo Engineer

  • 28 installs
  • 7 repo stars
  • Updated May 20, 2026
  • daemon-blockint-tech/agentic-enteprises-skill

Define and monitor SLA/SLO targets for service reliability.

About

SLA-SLO-engineer skill provides service level objective and agreement frameworks. Developers use it to define, monitor, and maintain service reliability targets.

  • SLA/SLO definition and monitoring
  • Error budget and reliability tracking

Sla Slo Engineer by the numbers

  • 28 all-time installs (skills.sh)
  • Ranked #872 of 1,435 DevOps & CI/CD skills by installs in the Skillselion catalog
  • Data as of Jul 29, 2026 (Skillselion catalog sync)
npx skills add https://github.com/daemon-blockint-tech/agentic-enteprises-skill --skill sla-slo-engineer

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs28
repo stars7
Last updatedMay 20, 2026
Repositorydaemon-blockint-tech/agentic-enteprises-skill

What it does

Define and monitor SLA/SLO targets for service reliability.

Files

SKILL.mdMarkdownGitHub ↗

SLA & SLO Engineer

When to Use

  • Select SLIs and document measurement queries, exclusions, and data sources
  • Set SLO targets, rolling windows, and per-journey or per-tier policies
  • Define error-budget math, consumption tracking, and policy actions (freeze, focus)
  • Design multi-window burn-rate alert policies and severity routing
  • Align customer-facing SLAs with internal SLOs (credits, measurements, carve-outs)
  • Tier services by criticality and map tiers to targets and review cadence
  • Run SLO review meetings, executive summaries, and quarterly governance
  • Estimate capacity headroom implied by latency or availability targets
  • Publish SLO specs for engineering (YAML/JSON schema, dashboard contracts)

When NOT to Use

  • Lead outage mitigation, paging, or on-call rotations → site-reliability-engineer, incident-management-engineer
  • Negotiate contract language, credits, or legal remedies → commercial-counsel
  • Build metrics/log/trace pipelines, collectors, or alertmanager config → devops, platform-engineer
  • Profile application code, load-test, or tune queries only → performance-engineer
  • Run production readiness reviews, chaos games, or release cutover → site-reliability-engineer, deployment-strategist
  • Coordinate multi-team program milestones without SLO scope → technical-program-manager

Related skills

NeedSkill
SRE execution: PRR, incident mitigation, chaos, release gatessite-reliability-engineer
Incident program, SEV, on-call, postmortemsincident-management-engineer
CI/CD pipelines, DORA, deploy gates wired to SLO policyci-cd-engineer
Delivery infra, GitOps, alert stack implementationdevops
IDP golden paths, platform SLOs for portal/scaffoldplatform-engineer
Load testing and latency profilingperformance-engineer
Rollout strategy and change tiersdeployment-strategist
Cross-team launch and RAIDtechnical-program-manager
Contractual SLA terms and redlinescommercial-counsel
Data pipeline freshness or warehouse SLAsdata-system-ops-lead

Core Workflows

1. Scope and principles

Service-level taxonomy, user-centric measurement, boundaries with SRE and legal.

See `references/sla_slo_scope_and_principles.md`.

2. SLI selection and measurement

Choose SLIs, define queries, exclusions, and validation.

See `references/sli_selection_and_measurement.md`.

3. SLO targeting and error budgets

Targets, windows, budget math, and policy actions.

See `references/slo_targeting_and_error_budgets.md`.

4. Alerting, burn rates, and policies

Multi-window alerts, routing, and noise control.

See `references/alerting_burn_rates_and_policies.md`.

5. Customer SLA vs internal SLO

Contract alignment, credits, carve-outs, and communication.

See `references/customer_sla_vs_internal_slo.md`.

6. Reporting, review, and governance

Cadences, dashboards, specs, and executive reporting.

See `references/reporting_review_and_governance.md`.

Outputs

  • SLO specification — SLI definition, query, target, window, exclusions, owners, tier
  • Error-budget policy — thresholds, actions, escalation, link to release policy
  • Burn-rate alert policy — windows, multipliers, severity, runbook links
  • SLA/SLO alignment matrix — customer metric ↔ internal SLI, measurement gaps, carve-outs
  • Tier catalog — criticality definitions with default targets and review cadence
  • SLO review pack — budget consumed, trends, top burners, proposed target changes
  • Capacity note — headroom vs latency/availability target (when in scope)

Principles

  • Measure user outcomes — availability and latency of journeys, not vanity infra metrics
  • Internal SLO stricter than external SLA — buffer for measurement lag and goodwill
  • Policy before panic — error-budget actions agreed before budget exhaustion
  • Alerts prove SLO risk — every page ties to budget burn or imminent breach
  • Govern with data — reviews change targets from evidence, not anecdotes
  • Hand off execution — SRE and IM own incident response; this skill owns the level definitions

When to load references

  • Scope and taxonomyreferences/sla_slo_scope_and_principles.md
  • SLI designreferences/sli_selection_and_measurement.md
  • Targets and budgetsreferences/slo_targeting_and_error_budgets.md
  • Burn alertsreferences/alerting_burn_rates_and_policies.md
  • Customer SLAreferences/customer_sla_vs_internal_slo.md
  • Reviews and governancereferences/reporting_review_and_governance.md

Related skills

DevOps & CI/CDmonitoring

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.