2000 character limit reached
Knowledge Return Oriented Prompting (KROP)
Published 11 Jun 2024 in cs.CR and cs.LG | (2406.11880v1)
Abstract: Many LLMs and LLM-powered apps deployed today use some form of prompt filter or alignment to protect their integrity. However, these measures aren't foolproof. This paper introduces KROP, a prompt injection technique capable of obfuscating prompt injection attacks, rendering them virtually undetectable to most of these security measures.
Paper Prompts
Sign up for free to create and run prompts on this paper.