Esc
There are no artifacts defined on this offensive technique (yet). Please consider contributing an addition to D3FEND.
LLM Prompt Self-Replication - AML.T0061
(ATLAS Technique)
Definition
An adversary may use a carefully crafted LLM Prompt Injection designed to cause the LLM to replicate the prompt as part of its output. This allows the prompt to propagate to other LLMs and persist on the system. The self-replicating prompt is typically paired with other malicious instructions (ex: LLM Jailbreak, LLM Data Leakage).
D3FEND Inferred Relationships
There are no artifacts defined on this offensive technique (yet). Please consider contributing an addition to D3FEND.