Catalog Home Page

Data representation influences protein secondary structure prediction using artificial neural networks

Lamont, O., Hiew, H.L. and Bellgard, M. (2001) Data representation influences protein secondary structure prediction using artificial neural networks. In: Seventh Australian and New Zealand Intelligent Information Systems Conference, 18 - 21 November 2001, Perth, Western Australia

[img]
Preview
PDF - Published Version
Download (437kB)

Abstract

Artificial Neural Networks (ANN) have been used very successfully for a number of classification problems in the molecular biology field. Protein secondary structure prediction is one of the oldest and best defined of these classification problems. Yet despite the considerable amount of work conducted in this field there still remain a number of fundamental computational issues that have not been thoroughly investigated, if considered at all. One important issue is identifying an appropriate data representation for input into the ANN. In this paper, we have investigated a range of new encoding schemes and evaluated their performance using recently introduced evaluation criterion. We have done this by preserving the redundant information of DNA codons that is lost when they are translated into amino acids. Interestingly, with our new data representation, the β-strand prediction performance was consistently higher (14% improvement) over the accuracy of the ANNs trained when the conventional representation was used.

Publication Type: Conference Paper
Murdoch Affiliation: School of Information Technology
Conference Website: http://ieeexplore.ieee.org/xpls/abs_all.jsp?arnumb...
URI: http://researchrepository.murdoch.edu.au/id/eprint/17331
Item Control Page Item Control Page

Downloads

Downloads per month over past year