Skip to main content

Integrating NOE and RDC using sum-of-squares relaxation for protein structure determination

Author(s): Khoo, Y; Singer, Amit; Cowburn, D

Download
To refer to this page use: http://arks.princeton.edu/ark:/88435/pr1d975
Full metadata record
DC FieldValueLanguage
dc.contributor.authorKhoo, Y-
dc.contributor.authorSinger, Amit-
dc.contributor.authorCowburn, D-
dc.date.accessioned2018-07-20T15:10:16Z-
dc.date.available2018-07-20T15:10:16Z-
dc.date.issued2017-07en_US
dc.identifier.citationKhoo, Y, Singer, A, Cowburn, D. (2017). Integrating NOE and RDC using sum-of-squares relaxation for protein structure determination. JOURNAL OF BIOMOLECULAR NMR, 68 (163 - 185. doi:10.1007/s10858-017-0108-7en_US
dc.identifier.issn0925-2738-
dc.identifier.urihttp://arks.princeton.edu/ark:/88435/pr1d975-
dc.description.abstractWe revisit the problem of protein structure determination from geometrical restraints from NMR, using convex optimization. It is well-known that the NP-hard distance geometry problem of determining atomic positions from pairwise distance restraints can be relaxed into a convex semidefinite program (SDP). However, often the NOE distance restraints are too imprecise and sparse for accurate structure determination. Residual dipolar coupling (RDC) measurements provide additional geometric information on the angles between atom-pair directions and axes of the principal-axis-frame. The optimization problem involving RDC is highly non-convex and requires a good initialization even within the simulated annealing framework. In this paper, we model the protein backbone as an articulated structure composed of rigid units. Determining the rotation of each rigid unit gives the full protein structure. We propose solving the non-convex optimization problems using the sum-of-squares (SOS) hierarchy, a hierarchy of convex relaxations with increasing complexity and approximation power. Unlike classical global optimization approaches, SOS optimization returns a certificate of optimality if the global optimum is found. Based on the SOS method, we proposed two algorithms-RDC-SOS and RDC-NOE-SOS, that have polynomial time complexity in the number of amino-acid residues and run efficiently on a standard desktop. In many instances, the proposed methods exactly recover the solution to the original non-convex optimization problem. To the best of our knowledge this is the first time SOS relaxation is introduced to solve non-convex optimization problems in structural biology. We further introduce a statistical tool, the Cram,r-Rao bound (CRB), to provide an information theoretic bound on the highest resolution one can hope to achieve when determining protein structure from noisy measurements using any unbiased estimator. Our simulation results show that when the RDC measurements are corrupted by Gaussian noise of realistic variance, both SOS based algorithms attain the CRB. We successfully apply our method in a divide-and-conquer fashion to determine the structure of ubiquitin from experimental NOE and RDC measurements obtained in two alignment media, achieving more accurate and faster reconstructions compared to the current state of the art.en_US
dc.format.extent163 - 185en_US
dc.language.isoen_USen_US
dc.relation.ispartofJOURNAL OF BIOMOLECULAR NMRen_US
dc.rightsAuthor's manuscripten_US
dc.titleIntegrating NOE and RDC using sum-of-squares relaxation for protein structure determinationen_US
dc.typeJournal Articleen_US
dc.identifier.doidoi:10.1007/s10858-017-0108-7-
dc.date.eissued2017-06-14en_US
dc.identifier.eissn1573-5001-
pu.type.symplectichttp://www.symplectic.co.uk/publications/atom-terms/1.0/journal-articleen_US

Files in This Item:
File Description SizeFormat 
1604.01504v5.pdf555 kBAdobe PDFView/Download


Items in OAR@Princeton are protected by copyright, with all rights reserved, unless otherwise indicated.