Open Internet by MindsNet
Vulnerability of AI defenses to multi-turn prompt injection attacks
Current AI defenses are ineffective against multi-turn prompt injection attacks, which can subtly influence a model over several interactions. Existing benchmarks only test one-shot attacks, leaving a significant research gap. This gap allows attackers to exploit AI models gradually, leading to potential security breaches.
Computing & Technology, Computer Science, Machine Learning