Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
datadog-labs avatar

Dd Monitors

  • 1.3k installs
  • 147 repo stars
  • Updated July 29, 2026
  • datadog-labs/agent-skills

dd-monitors provides documented workflows for Monitor management - list, search, file-based create, and alerting best practices.

About

The dd-monitors skill monitor management - list, search, file-based create, and alerting best practices. # Datadog Monitors Create, manage, and maintain monitors for alerting. ## Prerequisites This requires pup in your path. See [Setup Pup](https://github.com/datadog-labs/agent-skills/tree/main?tab=readme-ov-file#setup-pup). ## Command Execution Order (Token-Efficient) For scoped commands, use this order: 1. Check context first (prior outputs, conversation, saved values). If a required value is missing, run a discovery command first. If still ambiguous, ask the user to confirm. Then run the target command. Avoid speculative commands likely to fail. ## Quick Start ```bash pup auth login ``` ## Common Operations ### List Monitors ```bash pup monitors list pup monitors list --tags "team:platform" ``` ### Get Monitor ```bash pup monitors get <id> ``` ### Create Monitor ```bash pup monitors create --file monitor.json ``` ### Silence Alerts (Downtime) ```bash # No pup monitors mute/unmute commands. # Use downtime payloads to silence monitor notifications. pup downtime create --file downtime.json pup downtime cancel <downtime_id> ``` ## Monitor Creation Best Practices ### 1.

  • Check context first (prior outputs, conversation, saved values).
  • If a required value is missing, run a discovery command first.
  • If still ambiguous, ask the user to confirm.
  • Then run the target command.
  • Avoid speculative commands likely to fail.

Dd Monitors by the numbers

  • 1,258 all-time installs (skills.sh)
  • +51 installs in the week ending Aug 4, 2026 (Skillselion tracking)
  • Ranked #243 of 2,715 Automation & Workflows skills by installs in the Skillselion catalog
  • Security screen: LOW risk (skills.sh audit)
  • Data as of Aug 4, 2026 (Skillselion catalog sync)
At a glance

dd-monitors capabilities & compatibility

Capabilities
check context first (prior outputs, conversation · if a required value is missing, run a discovery · if still ambiguous, ask the user to confirm. · then run the target command. · avoid speculative commands likely to fail.
Use cases
documentation
From the docs

What dd-monitors says it does

# Datadog Monitors Create, manage, and maintain monitors for alerting.
SKILL.md
## Prerequisites This requires pup in your path.
SKILL.md
npx skills add https://github.com/datadog-labs/agent-skills --skill dd-monitors

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs1.3k
repo stars147
Security audit3 / 3 scanners passed
Last updatedJuly 29, 2026
Repositorydatadog-labs/agent-skills

How do I use dd-monitors for the task described in its SKILL.md triggers?

Monitor management - list, search, file-based create, and alerting best practices.

Who is it for?

Teams invoking dd-monitors when the user request matches documented triggers and prerequisites.

Skip if: Skip when cached docs are missing, the request is a negative trigger, or another sibling skill owns the workflow.

When should I use this skill?

Monitor management - list, search, file-based create, and alerting best practices.

What you get

Step-by-step guidance grounded in dd-monitors documentation and reference files.

  • datadog monitor YAML
  • silenced monitor state
  • pup CLI monitor listings

By the numbers

  • Skill version 1.0.1 from datadog-labs/agent-skills metadata
  • Matches monitor YAML via globs **/datadog*.yaml and **/*monitor*

Files

SKILL.mdMarkdownGitHub ↗

Datadog Monitors

Create, manage, and maintain monitors for alerting.

Prerequisites

This requires pup in your path. See Setup Pup.

Command Execution Order (Token-Efficient)

For scoped commands, use this order:

1. Check context first (prior outputs, conversation, saved values). 2. If a required value is missing, run a discovery command first. 3. If still ambiguous, ask the user to confirm. 4. Then run the target command. 5. Avoid speculative commands likely to fail.

Quick Start

pup auth login

Common Operations

List Monitors

pup monitors list
pup monitors list --tags "team:platform"

Get Monitor

pup monitors get <id>

Create Monitor

pup monitors create --file monitor.json

Silence Alerts (Downtime)

# No pup monitors mute/unmute commands.
# Use downtime payloads to silence monitor notifications.
pup downtime create --file downtime.json
pup downtime cancel <downtime_id>

Monitor Creation Best Practices

1. Avoid Alert Fatigue

RuleWhy
No flapping alertsUse last_Xm not last_1m
Meaningful thresholdsBased on SLOs, not guesses
Actionable alertsIf no action needed, don't alert
Include runbook@runbook-url in message
# WRONG - will flap constantly
query = "avg(last_1m):avg:system.cpu.user{*} > 50"  # ❌ Too sensitive

# CORRECT - stable alerting
query = "avg(last_5m):avg:system.cpu.user{env:prod} by {host} > 80"  # ✅ Reasonable window

2. Use Proper Scoping

# WRONG - alerts on everything
query = "avg(last_5m):avg:system.cpu.user{*} > 80"  # ❌ No scope

# CORRECT - scoped to what matters
query = "avg(last_5m):avg:system.cpu.user{env:prod,service:api} by {host} > 80"  # ✅

3. Set Recovery Thresholds

monitor = {
    "query": "avg(last_5m):avg:system.cpu.user{env:prod} > 80",
    "options": {
        "thresholds": {
            "critical": 80,
            "critical_recovery": 70,  # ✅ Prevents flapping
            "warning": 60,
            "warning_recovery": 50
        }
    }
}

4. Include Context in Messages

message = """
## High CPU Alert

Host: {{host.name}}
Current Value: {{value}}
Threshold: {{threshold}}

### Runbook
1. Check top processes: `ssh {{host.name}} 'top -bn1 | head -20'`
2. Check recent deploys
3. Scale if needed

@slack-ops @pagerduty-oncall
"""

NEVER Delete Monitors Directly

Use safe deletion workflow (same as dashboards):

def safe_mark_monitor_for_deletion(monitor_id: str, client) -> bool:
    """Mark monitor instead of deleting."""
    monitor = client.get_monitor(monitor_id)
    name = monitor.get("name", "")
    
    if "[MARKED FOR DELETION]" in name:
        print(f"Already marked: {name}")
        return False
    
    new_name = f"[MARKED FOR DELETION] {name}"
    client.update_monitor(monitor_id, {"name": new_name})
    print(f"✓ Marked: {new_name}")
    return True

Monitor Types

TypeUse Case
metric alertCPU, memory, custom metrics
query alertComplex metric queries
service checkAgent check status
event alertEvent stream patterns
log alertLog pattern matching
compositeCombine multiple monitors
apmAPM metrics

Audit Monitors

# Find monitors without owners
pup monitors list | jq '.[] | select(.tags | contains(["team:"]) | not) | {id, name}'

# Find noisy monitors (high alert count)
pup monitors list | jq 'sort_by(.overall_state_modified) | .[:10] | .[] | {id, name, status: .overall_state}'

Downtime vs Muting

UseWhen
DowntimeAny planned silence window
Monitor editQuery/threshold behavior changes
# Downtime (preferred)
pup downtime create --file downtime.json

Failure Handling

ProblemFix
Alert not firingCheck query returns data, thresholds
Too many alertsIncrease window, add recovery threshold
No data alertsCheck agent connectivity, metric exists
Auth errorpup auth refresh

References

Related skills

How it compares

Pick dd-monitors for Datadog-specific, YAML-driven monitor CRUD and silencing; use generic observability skills when you are not on Datadog or pup.

FAQ

What does dd-monitors do?

Monitor management - list, search, file-based create, and alerting best practices.

When should I use dd-monitors?

Monitor management - list, search, file-based create, and alerting best practices.

What are common prerequisites?

--- name: dd-monitors description: Monitor management - list, search, file-based create, and alerting best practices.

Is Dd Monitors safe to install?

skills.sh reports 3 of 3 security scanners passed. Review the Security Audits panel on this page before installing in production.

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.