Uncover prompt injection, insider threats with the Tenable One Model Refusal Detection
ID: a35040ce-9959-57ad-be6c-9edcebf5f479
STIX ID: report--a35040ce-9959-57ad-be6c-9edcebf5f479
Feed Name: Tenable Blog
This Tenable blog introduces Model Refusal Detection in Tenable One AI Exposure, which treats LLM refusals to risky prompts as early-warning signals to detect prompt injection, insider threats, and adversarial behavior. It summarizes research into refusal types, model variability, and advocates a defense-in-depth approach to monitor and investigate attempts to bypass model guardrails.
Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.
