logo

Block unsafe prompts targeting your LLM endpoints with Firewall for AI

ID: 3eebbe96-2df5-5612-b2d5-526ff535ed48

STIX ID: report--3eebbe96-2df5-5612-b2d5-526ff535ed48

Feed Name: Cloudflare Blog

Date Published: 2025-08-26

Date Updated: 2026-04-27

Author: Radwa Radwan

...
...

Cloudflare announced an expansion of its Firewall for AI to include unsafe content moderation using Llama Guard, enabling detection and blocking of harmful or sensitive LLM prompts (e.g., hate, violence, sexual content) at the network edge with unified analytics and policy enforcement. The system leverages an asynchronous, scalable architecture on Workers AI to minimize latency, exposes new analytics fields for unsafe topic insights, and allows custom rules to log or block categories, with a roadmap including prompt-injection/jailbreak detection and response handling; the feature is available in beta.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.