Skip to main navigation Skip to search Skip to main content

A study of quality measures for protein threading models

  • Susana Cristobal
  • , Adam Zemla
  • , Daniel Fischer
  • , Leszek Rychlewski
  • , Arne Elofsson
  • Uppsala University
  • Lawrence Livermore National Laboratory
  • International Institute of Molecular and Cell Biology in Warsaw
  • Stockholm University

Research output: Contribution to journalArticlepeer-review

191 Scopus citations

Abstract

Background: Prediction of protein structures is one of the fundamental challenges in biology today. To fully understand how well different prediction methods perform, it is necessary to use measures that evaluate their performance. Every two years, starting in 1994, the CASP (CriticalAssessment of protein Structure Prediction) process has been organized to evaluate the ability of different predictors to blindly predict the structure of proteins. To capture different features of the models, several measures have been developed during the CASP processes. However, these measures have not been examined in detail before. In an attempt to develop fully automatic measures that can be used in CASP, as well as in other type of benchmarking experiments, we have compared twenty-one measures. These measures include the measures used in CASP3 and CASP2 as well as have measures introduced later. We have studied their ability to distinguish between the better and worse models submitted to CASP3 and the correlation between them. Results: Using a small set of 1340 models for 23 different targets we show that most methods correlate with each other. Most pairs of measures show a correlation coefficient of about 0.5. The correlation is slightly higher for measures of similar types. We found that a significant problem when developing automatic measures is how to deal with proteins of different length. Also the comparisons between different measures is complicated as many measures are dependent on the size of the target. We show that the manual assessment can be reproduced to about 70% using automatic measures. Alignment independent measures, detects slightly more of the models with the correct fold, while alignment dependent measures agree better when selecting the best models for each target. Finally we show that using automatic measures would, to a large extent, reproduce the assessors ranking of the predictors at CASP3. Conclusions: We show that given a sufficient number of targets the manual and automatic measures would have given almost identical results at CASP3. If the intent is to reproduce the type of scoring done by the manual assessor in in CASP3, the best approach might be to use a combination of alignment independent and alignment dependent measures, as used in several recent studies.

Original languageEnglish
Article number5
JournalBMC Bioinformatics
Volume2
DOIs
StatePublished - Aug 1 2001

Fingerprint

Dive into the research topics of 'A study of quality measures for protein threading models'. Together they form a unique fingerprint.

Cite this