How Cloudflare built a multi-agent security operations harness for Managed Defense
Cloudflare describes an internal harness for its Managed Defense analysts that gathers evidence with deterministic code, filters noisy alerts with a triage model, and runs specialist AI agents in parallel to produce evidence-backed advisories. Analysts keep final responsibility for decisions, and the feature is in early beta.
The problem
Security alerts often arrive in bursts, and analysts must collect data, decide which alerts are related and account for missing sources while new ones keep coming. A first prototype that gave one general-purpose agent the whole investigation produced claims the evidence did not support, queried out-of-scope data, and hid lookup failures.
How they did it
- Run a fixed set of versioned reconnaissance workflows in application code before any model call, storing each datum with its source, version and timestamp.
- Score each alert against that recon data with Clef on Workers AI, and classify known high-volume noise as passive so it skips the active queue.
- Have a coordinator run four specialist agents in parallel: traffic analysis, customer context, global telemetry (aggregates only) and threat intelligence.
- Build a versioned evidence package, require specialists to cite items from it, and validate citations in application code.
- Let a synthesis agent combine the typed findings, with Clef picking from a reduced list of classifications, then generate an advisory report for the analyst.
- Orchestrate the stages with Workflows so a failed stage reuses already validated work, and record gaps as not checked, checked with no match, or checked with evidence of absence.
Results
- As reported by Cloudflare, the harness cuts the time analysts spend assembling and analyzing alerts, though the article gives no figures.
- Known high-volume noise is deterministically classified as passive so analysts are not paged repeatedly for it.
- The advisory shows related alerts, admitted evidence, visible gaps and recommended next steps, such as rate limiting or WAF rules.
- The early beta is available for eligible application-security alerts and cases in Cloudflare Managed Defense.
As reported by the source (Cloudflare Blog engineering post); AgentGid did not measure these figures.
This suits security teams with in-house engineers who can write versioned recon workflows and citation validators, and who already run on Cloudflare's stack (Workers AI, Workflows, Durable Objects, D1). Watch the evidence: the time savings are Cloudflare's own claim with no figures, and the feature is an early beta limited to eligible Managed Defense alerts. AgentGid has no listed alternatives for this case.
Similar use cases
Script Gemini CLI in your terminal: explain logs, write commits, document files
Gemini CLI's headless mode takes piped input and returns plain text or JSON. That makes it usable inside shell scripts, aliases and CI jobs.
Explanations of failures from piped logs, and commit messages generated from staged diffs.
Automate remediation after an AWS DevOps Agent investigation with Lambda Durable Functions
An AWS blog post shows how to pair AWS DevOps Agent, which only observes and reports, with a Lambda Durable Functions workflow that uses Amazon Bedrock to propose and apply fixes.…
In the demo, a function with a 3-second timeout was diagnosed, and Bedrock proposed raising it to 30 seconds.
How GitHub expanded secret validity checks with Copilot coding agent
GitHub's Secret Protection team had engineers do the research, then gave Copilot coding agent the repetitive work of adding validators for more leaked-token types. Coverage grew q…
As reported by GitHub: the team went from validating '32 partner token types' to onboarding 'almost 90 new types in just a few weeks'.