CORTEXA
← Browse
openalex˜The œinternational archives of the photogrammetry, remote sensing and spatial information sciences/International archives of the photogrammetry, remote sensing and spatial information sciences2026-07-23Cited by 0

The Emerging Role of Vision-Language Models in the Automation of Railway Asset Management: A Review and Future Perspective

Ashley Varghese, Mohammadjavad Ghorbanalivaki, Gunho Sohn

Abstract. The safety, efficiency, and longevity of global railway networks are directly linked to the rigorous inspection and management of their vast inventory of physical assets. Over the past decade, the field has progressed from manual surveys to automated systems leveraging imagery from track-based or aerial platforms. These systems predominantly built on traditional Computer Vision (CV) models have proven effective at detecting a pre-defined set of common assets. However, this progress has exposed a fundamental architectural and operational ceiling: the closed-world assumption. Current models are constrained to a fixed catalogue of classes defined during their training. It makes the model incapable of identifying novel objects or adapting to environmental changes without costly and continuous cycles of data re-annotation, retraining, and redeployment. This review paper argues that Vision-Language Models (VLMs), a paradigm whose rapid maturation is evidenced by recent comprehensive surveys offer a transformative solution. We provide a focused overview of the limitations of current CV systems and map the mechanics of a VLM-powered approach specifically Open-Vocabulary Detection and Reasoning Segmentation directly to the outstanding challenges in rail asset management. Ultimately, the literature suggests that the adoption of VLMs could catalyze a fundamental shift in railway infrastructure management that serves as a key enabler for next-generation Predictive Maintenance and autonomous Digital Twins.

View free PDFSource page

Related papers

openalex˜The œinternational archives of the photogrammetry, remote sensing and spatial information sciences/International archives of the photogrammetry, remote sensing and spatial information sciences2026-07-23

GeoOpen3D: Geometry-guided training-free open-vocabulary 3D segmentation via visual foundation models

Shuai Zhang, Zhuoxiao Li, Ou Jing, Tengxi Wang, Zhecheng Shi, Wufan Zhao

Abstract. Open-vocabulary 3D segmentation offers an attractive alternative to closed-set scene parsing, yet directly transferring 2D vision-language models to outdoor point clouds remains difficult because projection disrupts geometric continuity and sparse sampling weakens mask…

View free PDFSource page
openalex˜The œinternational archives of the photogrammetry, remote sensing and spatial information sciences/International archives of the photogrammetry, remote sensing and spatial information sciences2026-07-23

LLM-Supervised Point Cloud Processing: From Unsupervised 3D Scene-Graph Generation to Interactive Scene Manipulation

Florent Poux, Alex Key

Abstract. We demonstrate an end-to-end pipeline for 3D scene understanding which integrates unsupervised graph-based point cloud segmentation with LLM-enabled spatial reasoning and editing. A point cloud is segmented into a SemanticPatch decomposition (stage 1), labeled using a z…

View free PDFSource page
openalex˜The œinternational archives of the photogrammetry, remote sensing and spatial information sciences/International archives of the photogrammetry, remote sensing and spatial information sciences2026-07-23

BIM-to-Labelled Point Cloud: Automated Point Cloud Annotation from BIM Models using Bounding Boxes and Solid Geometries

Saad Boudarbala, Tania Landes, H. Macher, Thibault Bavoux

Abstract. This paper presents an automated framework for generating semantically labelled building point clouds from their corresponding BIM models. The proposed methodology aims to facilitate the creation of training datasets for deep learning–based indoor semantic segmentation.…

View free PDFSource page
openalex˜The œinternational archives of the photogrammetry, remote sensing and spatial information sciences/International archives of the photogrammetry, remote sensing and spatial information sciences2026-07-23

Enhancing Vision-Based Perception in Autonomous Driving: YOLO11–DETR Integration with Selection Model

Ahmed Reda, Naser El Sheimy, Adel Moussa

Abstract. Vision-based object detection is a key component of autonomous driving perception systems; however, models pretrained on large-scale generic datasets usually struggles when implemented in automotive environments due to domain shift. This research introduces a comprehens…

View free PDFSource page
openalex˜The œinternational archives of the photogrammetry, remote sensing and spatial information sciences/International archives of the photogrammetry, remote sensing and spatial information sciences2026-07-23

AI-Driven Extraction of Road Geometry and Asset Inventory from Mobile LiDAR Point Clouds

Divya Priya Balasubramani, Zaffar Sadiq Mohamed-Ghouse, Sanjay Khanna Diwakar, Ravichandran Narayanan, Muthu Kumara Samy Sudharsan

Abstract. The rapid urbanization and rising traffic volumes strain transportation infrastructure, demanding efficient road design auditing and asset management. Conventional manual surveys are labor-intensive and lack holistic three-dimensional context. This research presents an…

View free PDFSource page
openalex˜The œinternational archives of the photogrammetry, remote sensing and spatial information sciences/International archives of the photogrammetry, remote sensing and spatial information sciences2026-07-23

An Approach to 3D Digitisation and Segmentation of the Interior and Exterior of a Complex Museum Object

Simon Albers, Thomas Lühmann, Till Sieberth

Abstract. The digitisation of cultural heritage objects is an important procedure to conserve, share and analyse artefacts from the past. Nowadays, it is common practice to digitise artefacts using DSLR cameras and Structure from Motion. For most objects, this is a suitable proce…

View free PDFSource page