2003/10/20 by Wei‐Mou Zheng, Wei-Mou Zheng, Zheng, Wei-Mou
Biochemistry, Genetics and Molecular Biology · #Biomolecules (q-bio.BM) #FOS: Biological sciences #Machine Learning in Bioinformatics #Protein Structure and Dynamics #RNA and protein synthesis mechanisms #q-bio.BM
paper · pdf · doi:10.48550/arxiv.q-bio/0310026
8 pages, 1 figure, 2 tables
arxiv created 2003/10/20 · openalex publication_date 2003/10/20 · arxiv updated 2009/12/01 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
Instead of conformation states of single residues, refined conformation states of quintuplets are proposed to reflect conformation correlation. Simple hidden Markov models combining with sliding window scores are used for predicting secondary structure of a protein from its amino acid sequence. Since the length of protein conformation segments varies in a narrow range, we ignore the duration effect of the length distribution. The window scores for residues are a window version of the Chou-Fasman propensities estimated under an approximation of conditional independency. Different window widths are examined, and the optimal width is found to be 17. A high accuracy about 70% is achieved.