logo

Microsoft Unveils New AI Jailbreak That Allows Execution Of Malicious Instructions

ID: 2860c602-3e21-5b18-9423-37c706b3d3d3

STIX ID: report--2860c602-3e21-5b18-9423-37c706b3d3d3

Feed Name: cybersecurityNews.com

Threat Score
60/100

Date Published: 2024-06-27

Date Updated: 2026-04-21

Author: Tushar Subhra Dutta

...
...

Microsoft disclosed a new generative-AI jailbreak called “Skeleton Key,” a multi-step direct prompt-injection technique that can override responsible-AI guardrails and cause models to produce or follow harmful instructions. Microsoft tested base and hosted models from multiple vendors (Meta, Google, OpenAI, Mistral, Anthropic, Cohere) and rolled out mitigations including Prompt Shields and updates to Azure AI offerings; recommended mitigations include input filtering, system-message separation, output filtering, and abuse monitoring.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.