POODLE-L: a two-level SVM prediction system for reliably predicting long disordered regions

被引:107
作者
Hirose, Shuichi [1 ]
Shimizu, Kana
Kanai, Satoru
Kuroda, Yutaka
Noguchi, Tamotsu
机构
[1] PharmaDesign Inc, Tokyo 1040032, Japan
[2] Natl Inst Adv Ind Sci & Technol, Tokyo 1350064, Japan
[3] Tokyo Univ Agr & Technol, Grad Sch Engn, Dept Biotechnol & Life Sci, Koganei, Tokyo 1848588, Japan
关键词
D O I
10.1093/bioinformatics/btm302
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Motivation: Recent experimental and theoretical studies have revealed several proteins containing sequence segments that are unfolded under physiological conditions. These segments are called disordered regions. They are actively investigated because of their possible involvement in various biological processes, such as cell signaling, transcriptional and translational regulation. Additionally, disordered regions can represent a major obstacle to high-throughput proteome analysis and often need to be removed from experimental targets. The accurate prediction of long disordered regions is thus expected to provide annotations that are useful for a wide range of applications. Results: We developed Prediction Of Order and Disorder by machine LEarning (POODLE-L; L stands for long), the Support Vector Machines (SVMs) based method for predicting long disordered regions using 10 kinds of simple physico-chemical properties of amino acid. POODLE-L assembles the output of 10 two-level SVM predictors into a final prediction of disordered regions. The performance of POODLE-L for predicting long disordered regions, which exhibited a Matthew's correlation coefficient of 0.658, was the highest when compared with eight well-established publicly available disordered region predictors. Availability: POODLE-L is freely available at http://mbs.cbrc. jp/ poodle/poodle-l.html Contact: hirose-shuichi@aist.go.jp Supplementary information: Supplementary data are available at Bioinformatics online.
引用
收藏
页码:2046 / 2053
页数:8
相关论文
共 58 条
[11]   IUPred:: web server for the prediction of intrinsically unstructured regions of proteins based on estimated energy content [J].
Dosztányi, Z ;
Csizmok, V ;
Tompa, P ;
Simon, I .
BIOINFORMATICS, 2005, 21 (16) :3433-3434
[12]  
Dunker A K, 2000, Genome Inform Ser Workshop Genome Inform, V11, P161
[13]   Flexible nets - The roles of intrinsic disorder in protein interaction networks [J].
Dunker, AK ;
Cortese, MS ;
Romero, P ;
Iakoucheva, LM ;
Uversky, VN .
FEBS JOURNAL, 2005, 272 (20) :5129-5148
[14]   The protein trinity - linking function and disorder [J].
Dunker, AK ;
Obradovic, Z .
NATURE BIOTECHNOLOGY, 2001, 19 (09) :805-806
[15]  
Dunker AK, 2002, ADV PROTEIN CHEM, V62, P25
[16]   Intrinsic disorder and protein function [J].
Dunker, AK ;
Brown, CJ ;
Lawson, JD ;
Iakoucheva, LM ;
Obradovic, Z .
BIOCHEMISTRY, 2002, 41 (21) :6573-6582
[17]   Intrinsically disordered protein [J].
Dunker, AK ;
Lawson, JD ;
Brown, CJ ;
Williams, RM ;
Romero, P ;
Oh, JS ;
Oldfield, CJ ;
Campen, AM ;
Ratliff, CR ;
Hipps, KW ;
Ausio, J ;
Nissen, MS ;
Reeves, R ;
Kang, CH ;
Kissinger, CR ;
Bailey, RW ;
Griswold, MD ;
Chiu, M ;
Garner, EC ;
Obradovic, Z .
JOURNAL OF MOLECULAR GRAPHICS & MODELLING, 2001, 19 (01) :26-59
[18]   Intrinsically unstructured proteins and their functions [J].
Dyson, HJ ;
Wright, PE .
NATURE REVIEWS MOLECULAR CELL BIOLOGY, 2005, 6 (03) :197-208
[19]   Natively unfolded proteins [J].
Fink, AL .
CURRENT OPINION IN STRUCTURAL BIOLOGY, 2005, 15 (01) :35-41
[20]   Prediction of amyloidogenic and disordered regions in protein chains [J].
Galzitskaya, Oxana V. ;
Garbuzynskiy, Sergiy O. ;
Lobanov, Michail Yurievich .
PLOS COMPUTATIONAL BIOLOGY, 2006, 2 (12) :1639-1648