Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Liang Dai

Applying Deep Reinforcement Learning to the HP Model for Protein Structure Prediction

Dec 09, 2022

Kaiyuan Yang, Houjing Huang, Olafs Vandans, Adithya Murali, Fujia Tian, Roland H. C. Yap, Liang Dai

Figure 1 for Applying Deep Reinforcement Learning to the HP Model for Protein Structure Prediction

Figure 2 for Applying Deep Reinforcement Learning to the HP Model for Protein Structure Prediction

Figure 3 for Applying Deep Reinforcement Learning to the HP Model for Protein Structure Prediction

Figure 4 for Applying Deep Reinforcement Learning to the HP Model for Protein Structure Prediction

Abstract:A central problem in computational biophysics is protein structure prediction, i.e., finding the optimal folding of a given amino acid sequence. This problem has been studied in a classical abstract model, the HP model, where the protein is modeled as a sequence of H (hydrophobic) and P (polar) amino acids on a lattice. The objective is to find conformations maximizing H-H contacts. It is known that even in this reduced setting, the problem is intractable (NP-hard). In this work, we apply deep reinforcement learning (DRL) to the two-dimensional HP model. We can obtain the conformations of best known energies for benchmark HP sequences with lengths from 20 to 50. Our DRL is based on a deep Q-network (DQN). We find that a DQN based on long short-term memory (LSTM) architecture greatly enhances the RL learning ability and significantly improves the search process. DRL can sample the state space efficiently, without the need of manual heuristics. Experimentally we show that it can find multiple distinct best-known solutions per trial. This study demonstrates the effectiveness of deep reinforcement learning in the HP model for protein folding.

* Published at Physica A: Statistical Mechanics and its Applications, available online 7 December 2022. Extended abstract accepted by the Machine Learning and the Physical Sciences workshop, NeurIPS 2022

Via

Access Paper or Ask Questions

Sequential Monte Carlo Methods for System Identification

Mar 10, 2016

Thomas B. Schön, Fredrik Lindsten, Johan Dahlin, Johan Wågberg, Christian A. Naesseth, Andreas Svensson, Liang Dai

Figure 1 for Sequential Monte Carlo Methods for System Identification

Figure 2 for Sequential Monte Carlo Methods for System Identification

Abstract:One of the key challenges in identifying nonlinear and possibly non-Gaussian state space models (SSMs) is the intractability of estimating the system state. Sequential Monte Carlo (SMC) methods, such as the particle filter (introduced more than two decades ago), provide numerical solutions to the nonlinear state estimation problems arising in SSMs. When combined with additional identification techniques, these algorithms provide solid solutions to the nonlinear system identification problem. We describe two general strategies for creating such combinations and discuss why SMC is a natural tool for implementing these strategies.

* In proceedings of the 17th IFAC Symposium on System Identification (SYSID). Added cover page

Via

Access Paper or Ask Questions

Sparse Estimation From Noisy Observations of an Overdetermined Linear System

May 25, 2014

Liang Dai, Kristiaan Pelckmans

Figure 1 for Sparse Estimation From Noisy Observations of an Overdetermined Linear System

Figure 2 for Sparse Estimation From Noisy Observations of an Overdetermined Linear System

Figure 3 for Sparse Estimation From Noisy Observations of an Overdetermined Linear System

Figure 4 for Sparse Estimation From Noisy Observations of an Overdetermined Linear System

Abstract:This note studies a method for the efficient estimation of a finite number of unknown parameters from linear equations, which are perturbed by Gaussian noise. In case the unknown parameters have only few nonzero entries, the proposed estimator performs more efficiently than a traditional approach. The method consists of three steps: (1) a classical Least Squares Estimate (LSE), (2) the support is recovered through a Linear Programming (LP) optimization problem which can be computed using a soft-thresholding step, (3) a de-biasing step using a LSE on the estimated support set. The main contribution of this note is a formal derivation of an associated ORACLE property of the final estimate. That is, when the number of samples is large enough, the estimate is shown to equal the LSE based on the support of the {\em true} parameters.

* This paper is provisionally accepted by Automatica

Via

Access Paper or Ask Questions