GestaltMML: Enhancing Rare Genetic Disease Diagnosis through Multimodal Machine Learning Combining Facial Images and Clinical Texts
Authors:
Da Wu,
Jingye Yang,
Cong Liu,
Tzung-Chien Hsieh,
Elaine Marchi,
Justin Blair,
Peter Krawitz,
Chunhua Weng,
Wendy Chung,
Gholson J. Lyon,
Ian D. Krantz,
Jennifer M. Kalish,
Kai Wang
Abstract:
Individuals with suspected rare genetic disorders often undergo multiple clinical evaluations, imaging studies, laboratory tests and genetic tests, to find a possible answer over a prolonged period of time. Addressing this "diagnostic odyssey" thus has substantial clinical, psychosocial, and economic benefits. Many rare genetic diseases have distinctive facial features, which can be used by artifi…
▽ More
Individuals with suspected rare genetic disorders often undergo multiple clinical evaluations, imaging studies, laboratory tests and genetic tests, to find a possible answer over a prolonged period of time. Addressing this "diagnostic odyssey" thus has substantial clinical, psychosocial, and economic benefits. Many rare genetic diseases have distinctive facial features, which can be used by artificial intelligence algorithms to facilitate clinical diagnosis, in prioritizing candidate diseases to be further examined by lab tests or genetic assays, or in helping the phenotype-driven reinterpretation of genome/exome sequencing data. Existing methods using frontal facial photos were built on conventional Convolutional Neural Networks (CNNs), rely exclusively on facial images, and cannot capture non-facial phenotypic traits and demographic information essential for guiding accurate diagnoses. Here we introduce GestaltMML, a multimodal machine learning (MML) approach solely based on the Transformer architecture. It integrates facial images, demographic information (age, sex, ethnicity), and clinical notes (optionally, a list of Human Phenotype Ontology terms) to improve prediction accuracy. Furthermore, we also evaluated GestaltMML on a diverse range of datasets, including 528 diseases from the GestaltMatcher Database, several in-house datasets of Beckwith-Wiedemann syndrome (BWS, over-growth syndrome with distinct facial features), Sotos syndrome (overgrowth syndrome with overlapping features with BWS), NAA10-related neurodevelopmental syndrome, Cornelia de Lange syndrome (multiple malformation syndrome), and KBG syndrome (multiple malformation syndrome). Our results suggest that GestaltMML effectively incorporates multiple modalities of data, greatly narrowing candidate genetic diagnoses of rare diseases and may facilitate the reinterpretation of genome/exome sequencing data.
△ Less
Submitted 21 April, 2024; v1 submitted 23 December, 2023;
originally announced December 2023.
Locality of contacts determines the subdiffusion exponents in polymeric models of chromatin
Authors:
E. Marchi,
Y. Zhan,
G. Tiana
Abstract:
Loop extrusion by motor proteins mediates the attractive interactions in chromatin on the length scale of megabases, providing the polymer with a well-defined structure and at the same time determining its dynamics. The mean square displacement of chromatin loci varies from a Rouse-like scaling to a more constrained subdiffusion, depending on cell type, genomic region and time scale. With a simple…
▽ More
Loop extrusion by motor proteins mediates the attractive interactions in chromatin on the length scale of megabases, providing the polymer with a well-defined structure and at the same time determining its dynamics. The mean square displacement of chromatin loci varies from a Rouse-like scaling to a more constrained subdiffusion, depending on cell type, genomic region and time scale. With a simple polymeric model, we show that such a Rouse-like dynamics occurs when the parameters of the model are chosen so that contacts are local along the chain, while in presence of non-local contacts, we observe subdiffusion at short time scales with exponents smaller than 0.5. Such exponents are independent of the detailed choice of the parameters and build a master curve that depends only on the mean locality of the resulting contacts. We compare the loop-extrusion model with a polymeric model with static links, showing that also in this case only the presence of nonlocal contacts can produce low-exponent subdiffusion. We interpret these results in terms of a simple analytical model.
△ Less
Submitted 14 June, 2023;
originally announced June 2023.