Summary
Prompt injections, typically used by attackers to manipulate large language models, are now being adopted by cybersecurity defenders. Researchers from Tracebit found that placing prompt injections alongside sensitive data on Amazon Web Services can shut down AI hacking agents by directing LLMs to violate their guardrails.