Integrating NOE and RDC using sum-of-squares relaxation for protein structure determination
Author(s): Khoo, Y; Singer, Amit; Cowburn, D
DownloadTo refer to this page use:
http://arks.princeton.edu/ark:/88435/pr1d975
Abstract: | We revisit the problem of protein structure determination from geometrical restraints from NMR, using convex optimization. It is well-known that the NP-hard distance geometry problem of determining atomic positions from pairwise distance restraints can be relaxed into a convex semidefinite program (SDP). However, often the NOE distance restraints are too imprecise and sparse for accurate structure determination. Residual dipolar coupling (RDC) measurements provide additional geometric information on the angles between atom-pair directions and axes of the principal-axis-frame. The optimization problem involving RDC is highly non-convex and requires a good initialization even within the simulated annealing framework. In this paper, we model the protein backbone as an articulated structure composed of rigid units. Determining the rotation of each rigid unit gives the full protein structure. We propose solving the non-convex optimization problems using the sum-of-squares (SOS) hierarchy, a hierarchy of convex relaxations with increasing complexity and approximation power. Unlike classical global optimization approaches, SOS optimization returns a certificate of optimality if the global optimum is found. Based on the SOS method, we proposed two algorithms-RDC-SOS and RDC-NOE-SOS, that have polynomial time complexity in the number of amino-acid residues and run efficiently on a standard desktop. In many instances, the proposed methods exactly recover the solution to the original non-convex optimization problem. To the best of our knowledge this is the first time SOS relaxation is introduced to solve non-convex optimization problems in structural biology. We further introduce a statistical tool, the Cram,r-Rao bound (CRB), to provide an information theoretic bound on the highest resolution one can hope to achieve when determining protein structure from noisy measurements using any unbiased estimator. Our simulation results show that when the RDC measurements are corrupted by Gaussian noise of realistic variance, both SOS based algorithms attain the CRB. We successfully apply our method in a divide-and-conquer fashion to determine the structure of ubiquitin from experimental NOE and RDC measurements obtained in two alignment media, achieving more accurate and faster reconstructions compared to the current state of the art. |
Publication Date: | Jul-2017 |
Electronic Publication Date: | 14-Jun-2017 |
Citation: | Khoo, Y, Singer, A, Cowburn, D. (2017). Integrating NOE and RDC using sum-of-squares relaxation for protein structure determination. JOURNAL OF BIOMOLECULAR NMR, 68 (163 - 185. doi:10.1007/s10858-017-0108-7 |
DOI: | doi:10.1007/s10858-017-0108-7 |
ISSN: | 0925-2738 |
EISSN: | 1573-5001 |
Pages: | 163 - 185 |
Type of Material: | Journal Article |
Journal/Proceeding Title: | JOURNAL OF BIOMOLECULAR NMR |
Version: | Author's manuscript |
Items in OAR@Princeton are protected by copyright, with all rights reserved, unless otherwise indicated.