CORTEXA
← Browse
openalexElectronics2026-07-23Cited by 0

Benchmarking and Improving Perceptual Straightening for Robust AI-Generated Video Detection

Xenofon Papougiannakis, Rubén Tous

The proliferation of text-to-video generative models—including commercial systems such as Sora, Veo, Runway Gen-3, and Kling—makes automated detection of AI-generated videos an urgent problem in multimedia forensics. We present a study with two interleaved contributions: a cross-family benchmark and two improvements to the detection pipeline. For the benchmark, we evaluate three methodologically distinct detector families on three public datasets (VidProM, DeepAction, and DeCoF_T2V), spanning diverse open-source and commercial generators: ReStraV, a geometry-supervised detector grounded in the perceptual straightening hypothesis; D3, a training-free detector based on second-order XCLIP temporal statistics; and DeMamba, a supervised Bidirectional Mamba module on frozen XCLIP features. ReStraV and DeMamba achieve broadly comparable global performance on large datasets (VidProM AUC 0.985/0.989; DeCoF_T2V AUC 0.990/0.992), while D3 remains weak as a stand-alone detector under a fixed detection threshold calibrated on VidProM and transferred unchanged to the other benchmarks (AUC 0.534/0.410/0.454 on VidProM/DeepAction/DeCoF_T2V). Per-generator analysis reveals complementary failure modes: DeMamba recovers several hard commercial generators where ReStraV struggles, whereas ReStraV remains competitive in low-data settings at substantially lower computational cost. To bridge these weaknesses, we propose two improvements. First, Rich384 enriches ReStraV’s compact geometric descriptor with DINOv2 temporal embeddings, strengthening ranking quality on large benchmarks and recovering generators that geometry alone misses (AUC 0.995 on VidProM and 0.998 on DeCoF_T2V, with DeepAction AUC decreasing to 0.805). Second, we propose GurAI, our transparent logistic late-fusion method that combines Rich384 and DeMamba logits, raising threshold-free AUC to 0.998/0.999 on VidProM and DeCoF_T2V while preserving interpretability. On the small DeepAction cohort, fusion improves selected per-generator fake recalls at the cost of elevated REAL false positives (AUC 0.769; REAL accuracy 0.460).Our analysis suggests that transparent late fusion can exploit complementary detector strengths more effectively than architectural redesign alone when facing generator diversity.

View free PDFSource page

Related papers

crossrefElectronics2025-06-27Cited by 5

Machine Learning and Watermarking for Accurate Detection of AI-Generated Phishing Emails

Adrian Brissett, Julie Wall

Large Language Models offer transformative capabilities but also introduce growing cybersecurity risks, particularly through their use in generating realistic phishing emails. Detecting such content is critical; however, existing methods can be resource-intensive and slow to adap…

View free PDFSource page
crossrefElectronics2025-03-28Cited by 5

Novel Learning Framework with Generative AI X-Ray Images for Deep Neural Network-Based X-Ray Security Inspection of Prohibited Items Detection with You Only Look Once

Dongsik Kim, Jinho Kang

As the rapid expansion of future mobility systems increases, along with the demand for fast and accurate X-ray security inspections, deep neural network (DNN)-based systems have gained significant attention for detecting prohibited items by constructing high-quality datasets and…

View free PDFSource page
crossrefElectronics2025-12-06Cited by 1

A Dual Digital Twin Framework for Reinforcement Learning: Bridging Webots and MuJoCo with Generative AI and Alignment Strategies

Algirdas Laukaitis, Andrej Šareiko, Dalius Mažeika

Deep reinforcement learning (DRL) has shown potential for robotic training in virtual environments; however, challenges remain in bridging simulation and real-world deployment. This paper introduces an extended reinforcement learning framework that advances beyond traditional sin…

View free PDFSource page
crossrefElectronics2025-11-28Cited by 4

The Global Importance of Machine Learning-Based Wearables and Digital Twins for Rehabilitation: A Review of Data Collection, Security, Edge Intelligence, Federated Learning, and Generative AI

Maciej Piechowiak, Aleksander Goch, Ewelina Panas, Jolanta Masiak, Dariusz Mikołajewski, Izabela Rojek, et al.

The convergence of wearable technologies and digital twin (DT) systems is transforming rehabilitation engineering, enabling continuous monitoring, personalized therapeutic interventions, and predictive modeling of patient recovery pathways. This review examines the growing role o…

View free PDFSource page
crossrefElectronics2023-12-25Cited by 90

A Comprehensive Review of DeepFake Detection Using Advanced Machine Learning and Fusion Methods

Gourav Gupta, Kiran Raja, Manish Gupta, Tony Jan, Scott Thompson Whiteside, Mukesh Prasad

Recent advances in Generative Artificial Intelligence (AI) have increased the possibility of generating hyper-realistic DeepFake videos or images to cause serious harm to vulnerable children, individuals, and society at large with misinformation. To overcome this serious problem,…

View free PDFSource page
crossrefElectronics2023-10-17Cited by 1

Deep Learning Neural Network-Based Detection of Wafer Marking Character Recognition in Complex Backgrounds

Yufan Zhao, Jun Xie, Peiyu He

Wafer characters are used to record the transfer of important information in industrial production and inspection. Wafer character recognition is usually used in the traditional template matching method. However, the accuracy and robustness of the template matching method for det…

View free PDFSource page