CORTEXA
← Browse
zenodoPreprint2026-07-30

Fluent and Wrong: Eight Failure Patterns in Extended AI Use

Hector Ramos

Research on large language model failure predominantly assesses models in single exchanges and isolation: a model is prompted, the output is scored, and an error rate is reported. That is not how these systems are used in professional practice. Clinicians, attorneys, analysts, educators, and researchers work with them across extended sessions in which each output becomes the context for the next, and the final product carries a human signature. A test that scores a single response will not reveal these emerging behaviors. This paper describes eight failure patterns observed across extended sessions on multiple commercial AI platforms: recognition without correction, compounding rather than isolation of errors, halo effect from real anchors, unreliable self-explanation, degradation across long sessions, motivated framing under confrontation, sycophancy and confirmation bias amplification, and an architectural rather than statistical failure profile. Several are consistent with findings already established in the machine learning literature. Holistically, they are not addressed by lower hallucination rates, larger models, or vendor-side mitigations and instead are properties of system behavior under extended engagement, which carry direct consequences for verifying AI-assisted work in any domain where a person signs what the machine produced.

View free PDFSource page

Related papers

zenodoPreprint2026-07-29

MonteCarloJackknife.jl: Fast and Scalable Monte Carlo Approximation of Delete-d Jackknife Estimators in Julia

Soner AYDIN

This paper introduces MonteCarloJackknife.jl, an open-source Julia package that implements Monte Carlo approximation of delete-d jackknife estimators. Rather than exhaustively enumerating all deletion subsets, the package randomly samples a user-specified number of subsets, compu…

View free PDFSource page
zenodoPreprint2026-07-29

FOREX-SHIELD: A Multi-Modal Cyber-Defense Pipeline Combining Adversarially Hardened DeepLOB, Financial Transformers, and Zero-Knowledge Proofs for High-Frequency Foreign Exchange Settlement

Saiful Islam Tanvir

High-Frequency Foreign Exchange (FX) electronic execution networks process in excess of $7.5 trillion in daily spot volume across geographically distributed matching engines. Modern institutional trading infrastructure relies heavily on automated limit order book (LOB) forecastin…

View free PDFSource page
zenodoPreprint2026-07-28

Evaluation of complementary aspects of explainable AI techniques SHAP and LIME for deep neural networks for data sets in NLP domain

Ramesh Adeep Mohamed Arnest, Gursel Serpen

Abstract - Rapid advancements in large language models have enabled significant progress in solving complex real-world problems using deep neural networks (DNN). However, the black box nature of these DNN models poses significant challenges when it comes to explaining their decis…

View free PDFSource page