Datadog Alert Triage Agent
Live demo
Watch it work before it's live
[Triggered] api-gateway 5xx ratio above 2%
TriggeredError ratio 7.4% over the last 5 minutes against a 2% threshold, evaluated across 34 of 36 gateway pods. Triggered 14:07 UTC. Tags env:prod, region:eu-west-1, version:2026.8.4.
What you'll need
Set up in minutes
Connect Datadog
One sign-in. The agent acts through your account, scoped to what this template uses.
Connect Slack
One sign-in. The agent acts through your account, scoped to what this template uses.
Severity rules
Which monitors are worth waking someone for - what each severity level means on your services, what counts as fleet wide, your quiet hours, and the monitors you would rather never see in the channel.
Choose which agent runs it
NoClick's built-in models work out of the box — or bring Claude Code, Codex, and other coding agents on your own subscription.
Watch it handle a test run
A staged conversation against a simulated world — then it’s live.
The brief
About this agent
A Datadog page tells you a threshold moved and almost nothing else, so the first four minutes of every alert go on the same question: is this one sick host or is it everything. This agent answers that before a person opens a dashboard, reading the event and the monitor, counting what is actually alerting, pulling the metric around the window and any open incident on the service, then posting one short note to Slack. It suggests where to look next and touches none of your monitors, so the judgement stays with whoever is on call.
What people use it for
- One host or the fleet - The scope line comes first, with the host or tag count behind it. A single bad node reads very differently from thirty four pods, and that difference is what decides whether anyone gets out of bed.
- Numbers, not adjectives - The metric in the monitor's own query is pulled across the alert window and the hour before it, so the note carries real values and the times they were taken instead of the word elevated.
- Know if somebody already has it - Open incidents on the same service are checked before the message goes out, so the third alert of an ongoing outage arrives labelled as such rather than starting a parallel investigation in another channel.
- Quiet monitors stay quiet - Your severity rules decide what earns a channel post, so warning level noise and the monitors you have made peace with never reach Slack, and the ones that do are worth reading.
Before you fork
Which Datadog access does the agent need?
An API key and an application key that can read events, monitors, metrics, logs and incidents, plus the Slack channel you want the notes in. Nothing it does needs write scope, so a read only application key works in full. Your severity rules go in as a variable at setup.
Will it post on every warning we have configured?
Only on the monitors you route into the trigger, and then only when your severity rules say that level earns a channel post. Everything else is read, judged and dropped in silence. Start by pointing one noisy service at it for a day and read what it would have said.
Can it mute a monitor that is flapping at 3am?
No, and that is deliberate. It cannot mute, resolve or edit a monitor, or open or close an incident, so the most it can get wrong is a note you disagree with. Muting is a decision about what you are willing to miss, and it belongs to the person on call rather than to an agent reading a single event.
Keep exploring
More ways to run it
Run it with your coding agent
More agents like this
GitHub PR Review Agent
Reads every pull request the minute it opens and leaves one comment with a plain read of the change, a verdict on each…
Stripe Checkout Recovery Agent
Catches every abandoned Stripe checkout and leaves a short recovery email in Gmail naming the exact items and prices the…
Failed Payment Recovery Agent
Turns every failed charge into a ready-to-send dunning email in Gmail and one Slack line, with the amounts and dates pulled…
Stripe Revenue Digest Agent
Turns yesterday's Stripe activity into one Slack message every morning: what came in, what failed, who churned, and the one…
Make it yours
Put Datadog Alert Triage Agent to work.
Free to start. Guided setup, a test run against staged conversations, and it's live.