Automated discovery of functional components of proteins from amino-acid sequences based on rough sets and change of representation
Shusaku Tsumoto, Hiroshi Tanaka · 1995
Protein structure analysis from DNA sequences is an important and fast growing area in both computeT science and biochemistry. Although interesting ap-proaches have been studied, it is very dificult to cap-ture the characteristics of protein, since even a sim-ple protein are made of more than 100 amino acids, which makes biochemical experiments very dificult to detect functional components. For this reason, almost all the problems in this field are left unsolved and it is very important to develop a system which assists researchers on molecular biology to remove the difi-culties caused by combinatorial explosions. In this pa-per we report a system, called MWI (Molecular biol-ogists ’ Workbench version l.O), which extracts knowl-edge from amino-acid sequences by controlling appli-cation of domain knowledge automatically. We apply this method to comparative analysis of lysozyme and LX-lactalbumin. The results show that we obtain several interesting results from amino-acid sequences, which have not been reported before. 1.