Microsoft Unveils New AI Jailbreak That Allows Execution Of Malicious Instructions
ID: 2860c602-3e21-5b18-9423-37c706b3d3d3
STIX ID: report--2860c602-3e21-5b18-9423-37c706b3d3d3
Feed Name: cybersecurityNews.com
Microsoft disclosed a new generative-AI jailbreak called “Skeleton Key,” a multi-step direct prompt-injection technique that can override responsible-AI guardrails and cause models to produce or follow harmful instructions. Microsoft tested base and hosted models from multiple vendors (Meta, Google, OpenAI, Mistral, Anthropic, Cohere) and rolled out mitigations including Prompt Shields and updates to Azure AI offerings; recommended mitigations include input filtering, system-message separation, output filtering, and abuse monitoring.
Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.
