logo

ChatGPT Spills Secrets in Novel PoC Attack

ID: 2fc76ac6-7afd-541e-ac18-c1497bcbd699

STIX ID: report--2fc76ac6-7afd-541e-ac18-c1497bcbd699

Feed Name: Dark Reading

Threat Score
40/100

Date Published: 2024-03-13

Date Updated: 2026-04-21

Author: Jai Vijayan, Contributing Writer

...
...

Researchers from DeepMind, OpenAI, ETH Zurich, McGill and the University of Washington demonstrated a "top-down" attack that uses API queries to recover proprietary LLM architectural details (such as embedding projection matrices and hidden dimension sizes); they report recovering full projection matrices for some OpenAI models for under $20 and estimate other recoveries could cost under $2,000, raising concerns about model theft, reverse engineering, and reduced black‑box protections even though no active exploitation is reported.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.