Cloud & Deployment skills
Cloud services, infrastructure, deployment, and operations.
Skills in this category
55 skills in this collection
Agent Install
Install the Datadog Agent on Kubernetes using the Datadog Operator — required before enabling Single Step Instrumentation (SSI), which automatically instruments applications for APM without code changes. Only use if no Datadog Agent is…
Agent Observability Auto Experiment
Run an iterative code-improvement hill-climb against real Datadog LLM-Obs data, locally, with Claude Code as the agent. Establishes a baseline eval, makes one focused change, re-scores with the same harness, keeps the change if it…
Agent Observability Build Eval From Annotations
Fit a Datadog LLM-Obs evaluator to human labels. Takes an annotation queue, works out where in the trace the labelled property actually lives, drafts an LLM-judge that predicts the human label, scores that judge against the…
Agent Observability Eval Bootstrap
Bootstrap evaluators from production traces — by default propose online LLM-judge evaluators and, after you confirm, create them in Datadog as disabled drafts (never auto-enabled); on request emit Python SDK code or a framework-agnostic…
Agent Observability Eval Pipeline
End-to-end Agent Observability pipeline for an instrumented mlapp — classify production traces, root-cause failures, bootstrap evaluators, then (optionally) sample + publish a dataset, generate + run an experiment, and analyze results.…
Agent Observability Experiment Analyzer
Analyze LLM experiment results. Handles single or comparative experiments, exploratory or Q&A modes.
Agent Observability Experiment Bootstrap
Bootstrap a reproducible LLM Observability experiment through the Python ddtrace SDK or the Node dd-trace SDK. Use for experiment, dataset, evaluator, benchmark, regression, or LLM-as-a-judge scaffolding. The legacy Python invocation…
Agent Observability Replay Trace
Guidance for tasks where a developer wants to iterate on ONE specific Agent Observability / LLM Obs trace whose output they didn't like — re-running that trace against their LOCAL code, seeing a concise diff of the old vs new output, and…
Agent Observability Session Classify
Classify whether user intent was satisfied in a Datadog Agent Observability trace or session. Three modes: (1) sessionid — classify a single CMD+I assistant session with RUM; (2) traceid — classify a single Agent Observability trace…
Agent Observability Trace Rca
Root cause analysis on production LLM traces. Diagnoses why an LLM application is failing — works from eval judge verdicts, runtime errors, or structural anomalies depending on what signals are present. Walks the span tree from symptom to…
Agent Skills
Datadog skills for AI agents. Essential monitoring, logging, tracing and observability.
Azure Blob Storage for Python
Azure Blob Storage SDK for Python. Use for uploading, downloading, listing blobs, managing containers, and blob lifecycle.
Datadog App
Guides developers building Datadog Apps with TypeScript, React, the @datadog/apps scaffolder, and @datadog/vite-plugin.
Dd Account Setup
Ensure the user has an authenticated Datadog account with a valid DDAPIKEY on the right region before any Datadog setup or instrumentation. Detects existing DDAPIKEY / DDAPPKEY / DDSITE, validates them against the Datadog API, and fixes…
Dd Apm
APM - install, onboard, instrument, enable, set up, configure, traces, services, dependencies, performance analysis. Use for any request involving Datadog APM setup, instrumentation (SSI, ddtrace, agent install), or analysis.
Dd Audit
Audit Trail investigations - who changed what, key compromise, cost spike root cause, compliance evidence (SOC 2/PCI), and AI activity auditing.
Dd Audit AI Activity
Audit what the Bits AI assistant (MCP server) has done in your Datadog org — tool calls by user, resources accessed, and anomaly flags for AI governance.
Dd Audit Compliance Report
Generate auditor-ready compliance evidence from Datadog Audit Trail for SOC 2 and PCI DSS. Maps framework controls to specific query patterns and produces formatted output.
Dd Audit Cost Spike Investigation
Investigate a Datadog product usage or cost spike by correlating Usage Metering data (when/what spiked) with Audit Trail config changes (who changed what in the preceding window).
Dd Audit Key Compromise
Investigate a potentially compromised Datadog API key — timeline of actions, geo/IP breakdown, endpoints called, anomaly flags, and remediation steps.
20 of 55 skills loaded
Agent skills FAQ
An agent skill is a folder of instructions, usually headed by a SKILL.md file, that helps an AI assistant perform a specific task. A skill may also include scripts, examples, and reference files.