| Literature DB >> 33707455 |
Jordan T Ash1, Gregory Darnell2, Daniel Munro2, Barbara E Engelhardt3,4.
Abstract
Histopathological images are used to characterize complex phenotypes such as tumor stage. Our goal is to associate features of stained tissue images with high-dimensional genomic markers. We use convolutional autoencoders and sparse canonical correlation analysis (CCA) on paired histological images and bulk gene expression to identify subsets of genes whose expression levels in a tissue sample correlate with subsets of morphological features from the corresponding sample image. We apply our approach, ImageCCA, to two TCGA data sets, and find gene sets associated with the structure of the extracellular matrix and cell wall infrastructure, implicating uncharacterized genes in extracellular processes. We find sets of genes associated with specific cell types, including neuronal cells and cells of the immune system. We apply ImageCCA to the GTEx v6 data, and find image features that capture population variation in thyroid and in colon tissues associated with genetic variants (image morphology QTLs, or imQTLs), suggesting that genetic variation regulates population variation in tissue morphological traits.Entities:
Mesh:
Substances:
Year: 2021 PMID: 33707455 DOI: 10.1038/s41467-021-21727-x
Source DB: PubMed Journal: Nat Commun ISSN: 2041-1723 Impact factor: 14.919