CORTEXA
← Browse
crossrefMachine Learning: Science and Technology2026-07-06Cited by 0

Improving generalization and trainability of quantum eigensolvers via graph neural encoding

Jungyun Lee, Daniel K Park

Abstract Determining the ground state of a many-body Hamiltonian is a central problem across physics, chemistry, and combinatorial optimization, yet it is often classically intractable due to the exponential growth of Hilbert space with system size. Even on fault-tolerant quantum computers, quantum algorithms with convergence guarantees—such as quantum phase estimation and quantum subspace methods—require an initial state with sufficiently large overlap with the true ground state to be effective. Variational quantum eigensolvers (VQEs) are natural candidates for preparing such states; however, standard VQEs typically exhibit poor generalization, requiring retraining for each Hamiltonian instance, and often suffer from barren plateaus, where gradients can vanish exponentially with circuit depth and system size. To address these limitations, we propose an end-to-end representation learning framework that combines a graph autoencoder with a classical neural network to generate VQE parameters that generalize across Hamiltonian instances. By encoding interaction topology and coupling structure, the proposed model produces high-overlap initial states without instance-specific optimization. Through extensive numerical experiments on families of one- and two-local Hamiltonians, we demonstrate improved generalization and trainability, manifested as reduced test error and a significantly milder decay of gradient variance. We further show that our method substantially accelerates convergence in quantum subspace-based eigensolvers, highlighting its practical impact for downstream quantum algorithms.

View free PDFSource page

Related papers

crossrefMachine Learning: Science and Technology2026-04-27

A fully quantum-native recurrent neural network for end-to-end sequential learning on NISQ hardware

Rui Huang, Haibo Yi

Abstract Modeling temporal dependencies within quantum systems remains a key challenge for quantum machine learning. Current quantum neural networks largely depend on classical recurrent modules, which introduce optimization bottlenecks and coherence loss during sequence processi…

View free PDFSource page
crossrefMachine Learning: Science and Technology2026-05-21

Connectivity determines the capability of sparse neural network quantum states

Brandon Barton, Juan Carrasquilla, Christopher Roth, Agnes Valenti

Abstract The lottery ticket hypothesis (LTH) posits that within overparametrized neural networks, there exist sparse subnetworks that are capable of matching the performance of the original model when trained in isolation from the original initialization. We extend this hypothesi…

View free PDFSource page
crossrefMachine Learning: Science and Technology2026-07-02

Automatic charge state tuning of 300 mm silicon quantum dots using neural network segmentation of charge stability diagram

Peter Samaha, Amine Torki, Ysaline Renaud, Sam Fiette, Emmanuel Chanrion, Pierre-André Mortemousque, et al.

Abstract Tuning of gate-defined semiconductor quantum dots (QDs) is a major bottleneck for scaling spin-qubit technologies. We present a deep learning driven, semantic-segmentation pipeline that performs charge auto-tuning by locating transition lines in full charge stability dia…

View free PDFSource page
crossrefMachine Learning: Science and Technology2026-07-09

Data-driven surrogate modeling for thermal-hydraulic codes via hybrid deep neural networks and quantile learning

Hyojun Yi, Hyeonmin Kim, Seunghyoung Ryu

Abstract Nuclear energy is a clean, reliable power source, but realizing its potential requires strict safety measures in nuclear power plants. Thermal-hydraulic (TH) codes are used to simulate potential accident scenarios in probabilistic safety assessment (PSA). Their high comp…

View free PDFSource page
crossrefMachine Learning: Science and Technology2026-06-01

Convolutional neural network-driven preconditioners for conjugate gradients

Johannes Sappl, Viktor Daropoulos, Wolfgang Rauch, Matthias Harders

Abstract We present a data-driven approach for preconditioning large sparse symmetric positive definite linear systems using convolutional neural networks tailored for sparse tensor inputs. Our work targets system matrices arising from the discretization of partial differential e…

View free PDFSource page
crossrefMachine Learning: Science and Technology2026-05-28

NuGraph2 with explainability: post-hoc explanations for geometric neural network predictions

M Voetberg, Vitor F Grizzi, V Hewes, Giuseppe Cerati, Hadi Meidani

Abstract With the growing popularity of artificial intelligence (AI) used for scientific applications, the ability of attribute a result to a reasoning process from the network is in high demand for robust scientific generalizations to hold. In this work we aim to motivate the ne…

View free PDFSource page