Warning: Undefined array key "mm" in /www/wwwroot/www.ai-bt.com/si.php on line 10 Deprecated: trim(): Passing null to parameter #1 ($string) of type string is deprecated in /www/wwwroot/www.ai-bt.com/si.php on line 10 Discovering Collective Variables of Molecular Transitions via Genetic Algorithms and Neural Networks.

Literature DB >> 33662202

Discovering Collective Variables of Molecular Transitions via Genetic Algorithms and Neural Networks.

Ferry Hooft¹, Alberto Pérez de Alba Ortíz¹, Bernd Ensing¹.

Abstract

With the continual improvement of computing hardware and algorithms, simulations have become a powerful tool for understanding all sorts of (bio)molecular processes. To handle the large simulation data sets and to accelerate slow, activated transitions, a condensed set of descriptors, or collective variables (CVs), is needed to discern the relevant dynamics that describes the molecular process of interest. However, proposing an adequate set of CVs that can capture the intrinsic reaction coordinate of the molecular transition is often extremely difficult. Here, we present a framework to find an optimal set of CVs from a pool of candidates using a combination of artificial neural networks and genetic algorithms. The approach effectively replaces the encoder of an autoencoder network with genes to represent the latent space, i.e., the CVs. Given a selection of CVs as input, the network is trained to recover the atom coordinates underlying the CV values at points along the transition. The network performance is used as an estimator of the fitness of the input CVs. Two genetic algorithms optimize the CV selection and the neural network architecture. The successful retrieval of optimal CVs by this framework is illustrated at the hand of two case studies: the well-known conformational change in the alanine dipeptide molecule and the more intricate transition of a base pair in B-DNA from the classic Watson-Crick pairing to the alternative Hoogsteen pairing. Key advantages of our framework include the following: optimal interpretable CVs, avoiding costly calculation of committor or time-correlation functions, and automatic hyperparameter optimization. In addition, we show that applying a time-delay between the network input and output allows for enhanced selection of slow variables. Moreover, the network can also be used to generate molecular configurations of unexplored microstates, for example, for augmentation of the simulation data.

Entities: Chemical Disease Gene Species

Year: 2021 PMID： 33662202 DOI： 10.1021/acs.jctc.0c00981

Source DB: PubMed Journal: J Chem Theory Comput ISSN： 1549-9618 Impact factor: 6.006

2 in total

1. Sequence dependence of transient Hoogsteen base pairing in DNA.

Authors: Alberto Pérez de Alba Ortíz; Jocelyne Vreede; Bernd Ensing
Journal: PLoS Comput Biol Date: 2022-05-26 Impact factor: 4.779

Review 2. Collective variable-based enhanced sampling and machine learning.

Authors: Ming Chen
Journal: Eur Phys J B Date: 2021-10-20 Impact factor: 1.500

2 in total