CORTEXA
← Browse
arxivstat.MLcs.LG2026-07-08

Tensorized algorithms and scalable filtering methods for hidden Markov and factorial hidden Markov models

Roxana Barrios, Ioannis Sgouralis

A common method for the representation and analysis of time-series data is the hidden Markov model (HMM), where each observation is associated with a hidden state that evolves over time. However, many real-world systems are influenced by multiple independent factors, which are more naturally represented by factorial hidden Markov models (fHMM), where several hidden Markov chains jointly generate the observed data. Although an fHMM provides a richer and more realistic representation of many real-world systems, it can be reformulated as an equivalent HMM, but with a significantly larger state-space, leading to a severe increase in computational cost. In particular, the forward filtering algorithm, which is central to evaluation, decoding, and estimation tasks, becomes prohibitively expensive even for small systems. This work focuses on developing scalable methods for time-series analysis using tensor algebra to exploit the multidimensional structure of fHMM directly, without constructing intermediate HMM representations. Our novel filtering approach significantly improves computational performance and enables the efficient analysis of large systems and datasets, extending the scope of fHMM and providing a practical framework for data intensive applications.

View free PDFSource page

Related papers

arxivcs.LGstat.ML2026-07-10

Estimation, Prediction, and Assortment Optimization for Markov Chain Choice Models with Panel Data

Yalcin Akcay, Gerardo Berbeglia, Young-San Lin

We propose a framework for the Markov chain (MC) choice model with panel data, including parameter estimation, personalized choice prediction, and personalized assortment optimization. In contrast to the traditional setting, which assumes that each transaction is independently dr…

View free PDFSource page
arxivstat.MLcs.LGstat.ME2026-07-20

An efficient adaptive dimension selection algorithm for multidimensional probit graded response models

Yu Zhou, Yincai Tang, Bin Lv, Meng Gao

Multidimensional graded response models (MGRMs) are widely used for analyzing ordinal questionnaire data in psychological and educational assessments. A central challenge in applying these models is determining the number of latent dimensions. Conventional approaches usually fit…

View free PDFSource page
arxivstat.MLcs.LGmath.NAstat.CO2026-07-24

Convergence analysis of a family of Zermelo-type iterations for the Bradley--Terry model

Ruijian Han, Ding Lu, Yiming Xu

Zermelo's algorithm is a classical method for computing the maximum likelihood estimator in the Bradley--Terry (BT) model, but its convergence can be slow in practice. To accelerate computation, Newman introduced a family of Zermelo-type fixed-point iterations parameterized by $α…

View free PDFSource page
arxivstat.MLcs.LGstat.ME2026-07-14

LatentFlow: A General Framework for Conditioning Stochastic Processes

Louis Sharrock, Lachlan Astfalck, Henry Moss

Stochastic-process models are, as a rule, far easier to simulate than to condition. Non-linear observations, non-Gaussian likelihoods, black-box information, and global constraints all induce intractable conditional laws, requiring bespoke, model-specific constructions. We introd…

View free PDFSource page
arxivstat.MLcs.LGstat.CO2026-07-06

Integrating Neural Encoders in Bayesian Generalized Linear Mixed Models for Multimodal Data

Yuankang Zhao, Youngsoo Baek, Felipe A. Medeiros, Samuel Berchuck, Matthew M. Engelhard

Scalable Bayesian inference for generalized linear mixed models (GLMMs) provides uncertainty-aware analysis of correlated longitudinal data, but existing scalable approaches largely assume low-dimensional tabular predictors and do not directly accommodate high-dimensional modalit…

View free PDFSource page