CORTEXA
← Browse
crossrefMachine Learning and Knowledge Extraction2026-06-18Cited by 0

XAI2Brain: A Perspective on Mechanistic Interpretability for Brain–AI Alignment

Richard Jiang, Yongchen Zhou, Boyuan Wang, Plamen Angelov, Qiang Ni

The convergence of artificial intelligence (AI), explainable AI (XAI), and neuroscience is fostering new opportunities for understanding both machine and biological intelligence through interpretable and human-centered learning paradigms. In this Perspective, we introduce XAI2Brain as a conceptual framework for brain–AI alignment, positioning mechanistic interpretability as an intermediate layer connecting neural network representations, human understanding, and neuroscience-inspired AI design. Rather than viewing XAI solely as a post hoc transparency tool, we emphasize its emerging role in enabling mechanistic analysis of internal model representations, concept-level reasoning, and interactive human–AI alignment. We define XAI2Brain as a multi-level conceptual framework rather than a deployable system, explicitly aimed at structuring brain–AI alignment across representation-level, mechanism-level, and interaction-level perspectives. We survey the evolution of XAI methodologies—from feature attribution and concept-based explanations to mechanistic and human-centric interpretability approaches—and discuss how these methods may support bidirectional knowledge transfer between AI systems and cognitive neuroscience. Importantly, we adopt a cautious stance on brain–AI analogy, explicitly recognizing that artificial neural representations are not equivalent to biological neural representations, and instead focusing on functional and informational correspondences rather than structural equivalence. Unlike conventional human-in-the-loop or reinforcement learning from human feedback paradigms that primarily optimize behavioral outputs, XAI2Brain focuses on cognitively interpretable and mechanistically grounded alignment between AI systems and human reasoning processes. This alignment promotes interactive human-in-the-loop intelligence, empowering humans to comprehend, guide, and refine AI systems, while enabling AI systems to better interpret human instructions, intentions, and contextual reasoning. We further discuss the challenges of scaling explainability to large generative and multimodal models, including issues of interpretability robustness, cognitive compatibility, evaluation, and ethical accountability. We also highlight key limitations of current mechanistic interpretability methods, including explanation instability, representation superposition, and lack of causal guarantees, underscoring that these challenges remain open research problems. Rather than proposing a complete artificial brain architecture, this Perspective outlines a research roadmap toward more interpretable, adaptive, and neuroscience-inspired AI systems capable of supporting future brain–AI integration and collaborative intelligence. We additionally clarify that this work follows a narrative perspective review methodology with structured thematic synthesis of the literature. By framing explainability as a bridge between mechanistic AI understanding, cognitive science, and human-centered interaction, XAI2Brain highlights the importance of interpretable alignment for the next generation of brain-inspired AI systems.

View free PDFSource page

Related papers

crossrefMachine Learning and Knowledge Extraction2026-07-23

From Black Box to Clarity: A Systematic Review of Explainability Methods in Deep Convolutional Neural Networks

Zina Tayari, Mourad Zaied

Deep neural networks (DNNs) have significantly advanced machine perception and reasoning; however, their lack of transparency in decision-making continues to pose a major challenge, particularly in high-stakes domains such as healthcare, finance, and law. This is especially conce…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2026-07-22

Rapid Machine Learning–Driven Modeling for Large-Scale Validation and Optimization of Control Variables in Wireless Power Transfer Systems

Oscar García-Izquierdo, José Francisco Sanz, Juan Luis Villa, María Paz Comech, Julio J. Melero

Validating wireless power transfer (WPT) systems for electric vehicles (EVs) is a challenge due to efficiency variations caused by coil misalignments and height differences arising from various vehicle designs. Traditional simulation methods, such as finite element analysis (FEM)…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2026-07-22

Alzheimer’s Disease Detection Based on Machine Learning and Deep Learning Frameworks: A Cross-Dataset Comparative Performance Analysis and Assessment of Clinical Readiness

Keenan Ramnarain, Rito Clifford Maswanganyi, Philani Khumalo

Alzheimer’s disease (AD) is the most prevalent neurodegenerative disorder worldwide, affecting approximately 56.9 million people in 2021 and projected to reach 152 million by 2050. Its defining pathological features, amyloid-beta plaques and neurofibrillary tangles, accumulate fo…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2026-07-22

Cognitive Friction in Clinical Decision Support: A Comparative Study of Judicial and Adjunct Human–AI Interaction Protocols

Samuele Pe, Laura Bergomi, Giovanna Nicora, Camilla A. Simonelli, Prabhjot Kour, Esperanza Diaz, et al.

Artificial intelligence is increasingly used to support clinical decision making, yet concerns remain regarding algorithmic aversion, automation bias and the preservation of meaningful human oversight; while explainable AI aims to improve transparency, less attention has been dev…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2026-07-16

From Assets and Processes to Service Ecosystems: A Hierarchical Digital Twin Framework for Knowledge Representation

Igor Kabashkin

Digital twins (DTs) have become a central paradigm for modeling cyber–physical systems and digital infrastructures, yet the term is applied to very different representations—from physical assets to operational processes and service environments. This ambiguity obscures how the va…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2026-07-15

Beyond Forecast Accuracy: Evaluating the Error–Profit Paradox in AI-Based Copper Price Prediction

László Vancsura, Tibor Tatay, Tibor Bareith

Copper is a strategically important commodity whose price dynamics are increasingly affected by structural changes, geopolitical shocks, and the global energy transition. These conditions create substantial challenges for forecasting models and provide a useful setting for evalua…

View free PDFSource page