CORTEXA
← Browse
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24Cited by 0

PromptShield AI: A Multi-Agent Architecture for Intelligent Prompt Injection and Jailbreak Attack Detection Using Machine Learning

Jahnavi Somaraju, N. Sree Charan, M. Mythili, T. Reddy Bhargavi, K. Navya Sree

Large language models (LLMs) are increasingly deployed in user-facing applications, which exposes them to prompt injection and jailbreak attacks that override system instructions, exfiltrate data, or elicit disallowed behaviour. Existing defences are largely single-mechanism: a rule filter, a single machine-learning (ML) classifier, or a single LLM-based judge, each of which is comparatively easy to evade once its decision boundary is known. This paper proposes PromptShield AI, a multi-agent architecture that fuses lexical, statistical, semantic, and contextual detection signals through a coordinated set of specialised agents rather than a single monolithic classifier. An Orchestrator Agent decomposes each incoming request and routes it to five parallel detection agents; a Decision Fusion Agent calibrates and combines their outputs; a Response/Mitigation Agent executes the resulting policy (allow, sanitize, or block); and a Logging and Feedback Agent closes the loop for continuous retraining. We describe the architecture, the inter-agent communication protocol, the fusion algorithm, and a reference implementation, and we outline an experimental protocol together with illustrative evaluation results and an ablation study. The results indicate that the coordinated multi-agent design improves detection F1-score and reduces false-positive rate relative to any individual detection mechanism, at an acceptable latency overhead, while remaining more resilient to obfuscation and multi-turn attacks than single-agent baselines.

View free PDFSource page

Related papers

openalexZenodo (CERN European Organization for Nuclear Research)2026-07-25

Multi-Modal Deepfake Detection System Using Hybrid Deep Learning on Visual and Audio Features

Nandana K Gowda, P Hemavathi

Abstract: Deepfake technology, driven by generative models such as GANs and diffusion architectures, has enabled the creation of highly realistic manipulated media capable of deceiving both visual and auditory perception. Such forgeries pose significant risks to identity verifica…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-25

Event-Based Prediction of Liquidity Sweep Dynamics in XAUUSD Using Machine Learning

Vanshvardhan Sharma

This paper develops a machine learning framework for detecting and predicting liquidity sweep events in XAUUSD using event-based market microstructure analysis. Using 15-minute data from 2014–2024, the study formalizes liquidity sweeps as a binary classification problem evaluated…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24

SkillGraph: A Multi-Agent Architecture for AI-Powered Career Recommendation Using Knowledge Graphs and Graph Neural Networks

Jahnavi Somaraju, V. Guru Thrinath, S. R. Bhavishya, S. Bhavya, Y. Jahnavi

The rapid growth of online career and learning resources has made it difficult for job seekers and professionals to identify the skills, roles, and learning paths that best match their goals. This paper presents SkillGraph, a multi-agent architecture for AIpowered career recommen…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-25

Suspicious Activity Detection Using Machine Learning

Jyoti Neeli, S V Shwetha, H M Rakshitha, Nuthan

Abstract: Video surveillance has become a critical component in today's world. With developments in advanced systems have been developed as a result of the precision and efficacy of deep learning, machine learning, and artificial intelligence to identify and identify questionable…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-07-24

Cloud-Native Adversarial Machine Learning: A Serverless Cybersecurity Architecture for Neutralizing Prompt Injections in Large Language Models

YINKA ADERIBIGBE

The integration of Large Language Models into enterprise network architectures has introduced severe cybersecurity vulnerabilities, most notably adversarial prompt injection and zero-day data extraction attacks. Traditional network security protocols are fundamentally ill-equippe…

View free PDFSource page
openalexZenodo (CERN European Organization for Nuclear Research)2026-08-09

A Systematic Review of Machine Learning, Deep Learning, and Explainable AI Approaches for Cardiac Disease Prediction

Sunanda Budihal, Sheetalrani Kawale, Abhishek Angadi

The cardiovascular (Cardiac) disease (CVD) is another factor that causes death among the global population most, and this is the reason why there is a high necessity to implement proper, effective, and interpretive diagnostic systems. The usage of machine learning (ML), deep lear…

Also available via: European Organization for Nuclear Research

View free PDFSource page