Agent security & threat detection

Azure AI Content Safety Prompt Shields

Microsoft API in Azure AI Content Safety that analyses user prompts and supplied documents for adversarial instructions, returning attack-detected flags for direct user prompt attacks and indirect attacks embedded in external content. It is a detection endpoint that applications must act on rather than an enforcement gateway.

commercial · generally available · Research snapshot 2026-09-06

Visit the official product source ↗

Where it fits

Agent security & threat detection

Useful conversation with: AI application developer, Cloud security architect, Responsible AI lead.

Ask for a demonstration

Show me a shieldPrompt call flagging an indirect injection hidden in an uploaded document, and how the calling agent enforces a block based on that response.

Capabilities and evidence

Support labels reflect the supplied research. Documentation and vendor claims are not independent product tests. “Not established” means the researcher did not find support; it does not prove a capability is absent.

Documented by provider

Prompt Shields is a unified API that detects and blocks adversarial user input attacks on LLMs, covering user prompt injection attempts to circumvent system rules and document-based indirect attacks with hidden instructions in external material.

Limit: Detection is scoped to prompts and documents; tool calls, agent plans and outbound data flows are not covered.

Source s1

Documented by provider

The API endpoint contentsafety/text:shieldPrompt returns attackDetected booleans for the user prompt and each document, with configurable thresholds for filtering.

Limit: Enforcement depends on the calling application; the API itself does not block traffic.

Source s2

Limitations to discuss

Sources

  1. Prompt Shields in Azure AI Content Safety · Microsoft · official docs
    Access date reported by researcher: 2026-09-06
  2. Quickstart: Detect prompt attacks with Prompt Shields · Microsoft · official docs
    Access date reported by researcher: 2026-09-06

Listing does not imply partnership, supplier status, a working DutyGraph integration, or a compliance certification.

Suggest a correction