Integrating NOE and RDC using sum-of-squares relaxation for protein structure determination

Y. Khoo, A. Singer, D. Cowburn

Research output: Contribution to journalArticlepeer-review

3 Scopus citations

Abstract

We revisit the problem of protein structure determination from geometrical restraints from NMR, using convex optimization. It is well-known that the NP-hard distance geometry problem of determining atomic positions from pairwise distance restraints can be relaxed into a convex semidefinite program (SDP). However, often the NOE distance restraints are too imprecise and sparse for accurate structure determination. Residual dipolar coupling (RDC) measurements provide additional geometric information on the angles between atom-pair directions and axes of the principal-axis-frame. The optimization problem involving RDC is highly non-convex and requires a good initialization even within the simulated annealing framework. In this paper, we model the protein backbone as an articulated structure composed of rigid units. Determining the rotation of each rigid unit gives the full protein structure. We propose solving the non-convex optimization problems using the sum-of-squares (SOS) hierarchy, a hierarchy of convex relaxations with increasing complexity and approximation power. Unlike classical global optimization approaches, SOS optimization returns a certificate of optimality if the global optimum is found. Based on the SOS method, we proposed two algorithms—RDC-SOS and RDC–NOE-SOS, that have polynomial time complexity in the number of amino-acid residues and run efficiently on a standard desktop. In many instances, the proposed methods exactly recover the solution to the original non-convex optimization problem. To the best of our knowledge this is the first time SOS relaxation is introduced to solve non-convex optimization problems in structural biology. We further introduce a statistical tool, the Cramér–Rao bound (CRB), to provide an information theoretic bound on the highest resolution one can hope to achieve when determining protein structure from noisy measurements using any unbiased estimator. Our simulation results show that when the RDC measurements are corrupted by Gaussian noise of realistic variance, both SOS based algorithms attain the CRB. We successfully apply our method in a divide-and-conquer fashion to determine the structure of ubiquitin from experimental NOE and RDC measurements obtained in two alignment media, achieving more accurate and faster reconstructions compared to the current state of the art.

Original languageEnglish (US)
Pages (from-to)163-185
Number of pages23
JournalJournal of Biomolecular NMR
Volume68
Issue number3
DOIs
StatePublished - Jul 1 2017

All Science Journal Classification (ASJC) codes

  • Biochemistry
  • Spectroscopy

Keywords

  • Convex optimization
  • Cramér–Rao lower-bound
  • Nuclear Overhauser effect
  • Protein structure determination
  • Residual dipolar coupling
  • Semidefinite programming
  • Sum-of-squares optimization

Fingerprint

Dive into the research topics of 'Integrating NOE and RDC using sum-of-squares relaxation for protein structure determination'. Together they form a unique fingerprint.

Cite this