openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24
PromptShield AI: A Multi-Agent Architecture for Intelligent Prompt Injection and Jailbreak Attack Detection Using Machine Learning
Jahnavi Somaraju, N. Sree Charan, M. Mythili, T. Reddy Bhargavi, K. Navya Sree
Large language models (LLMs) are increasingly deployed in user-facing applications, which exposes them to prompt injection and jailbreak attacks that override system instructions, exfiltrate data, or elicit disallowed behaviour. Existing defences are largely single-mechanism: a r…