logo

Anthropic Research

ID: b6040af0-1bcc-5815-a08f-b6ee8cf87625

STIX ID: identity--b6040af0-1bcc-5815-a08f-b6ee8cf87625

Feed Type: skeleton

Earliest post: 2023-11-03

Latest post: 2026-07-28

Anthropic’s research explores AI safety, model interpretability, economic impact, and emerging risks, with the goal of making increasingly capable AI systems helpful, honest, and safe.

01/01/2020
08/11/2026
Title Date Published Describes IncidentAuthorVisible
Discovering cryptographic weaknesses with Claude2026-07-23TrueTrue
Measuring LLMs’ ability to develop exploits2026-06-05TrueTrue
Assessing Claude Mythos Preview’s cybersecurity capabilities2026-04-16TrueTrue
AI agents find smart contract exploits2026-03-24TrueTrue
Cyber toolkits for LLMs2026-03-24TrueTrue
Reverse engineering Claude's CVE-2026-2796 exploit2026-03-24TrueTrue
AI models on realistic cyber ranges2024-11-02TrueTrue
Many-shot jailbreaking2024-03-29TrueTrue
A Mathematical Framework for Transformer Circuits2023-12-18TrueTrue
Measuring LLMs’ impact on N-day exploits2023-11-03TrueTrue
A small number of samples can poison LLMs of any size2023-11-03TrueTrue
Mapping AI-enabled cyber threats2023-11-03TrueTrue
Project Glasswing: An initial update2023-11-03TrueTrue
Mitigating the risk of prompt injections in browser use2023-11-03TrueTrue
LLM-discovered 0 days2023-11-03TrueTrue

1–15 of 15