Back

Quantum Deep Learning Pipeline for Next Generation Network Biology

Yazdi, M.; Srinivas, K. R.; Yadav, A.

2025-10-29 bioinformatics
10.1101/2025.10.28.685074 bioRxiv
Show abstract

PurposeModule discovery in omics networks is central to interpretation. Classical pipelines capture broad community structure, but exact search for small, connected, topology aware modules is combinatorial, making exhaustive solutions impractical at genome scale. Moreover, most quantum clique formulations to date optimize only maximal density on non biological graphs, which mismatches the heterogeneous shapes of real biological modules. MethodsWe trained a symmetric 10-dimensional autoencoder on GTEx normals (Heart Left Ventricle; Muscle Skeletal; UCSC Xena Toil) to obtain tissue-specific latent representations. For each tissue we built a 10*10 latent-node correlation graph and formulated a QUBO that rewards strong edges and penalizes isolated selections (edge threshold {tau}=0.80). QAOA (depth=3, COBYLA) generated high-probability bitstrings, which we post-filtered to retain connected, non-overlapping subsets; decoder triggered gene sets and multi-library enrichment provided functional interpretation. ResultsQAOA distilled the Heart graph to three dyads (2-node modules with strong-edge sums {approx}0.89/0.95), pairing conduction with mitochondria, conduction with ECM/adhesion, and two contractile nodes. In Muscle, QAOA returned one dominant 6-node module (edge-sum {approx}7.0) integrating contractile machinery, electrophysiology, mitochondrial metabolism, and translation; a weaker ECM leaning triplet was visible but fell below the top 10 threshold. ConclusionQAOA based quantum optimization yields discrete, testable modules from latent correlations that match tissue programs. Despite a 10-node demo (current limits), the scale-agnostic design extends to 10^3 10^5 nodes and multi omics via hierarchical compression and hybrid search supporting quantum modeling for next generation network biology.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.