Article
Application of a novel and fast information-theoretic method to the discovery of higher-order correlations in protein databases.
Pacific Symposium on Biocomputing. Pacific Symposium on Biocomputing - 1 Jan 1998
Steeg E W, Pham H
Abstract excerpt
We present a fast, discrete data-mining approach to the problem of finding kappa-tuples of correlated amino acid residues in protein sequence data. When sets of sequence-distant sites display high mutual information, they may bespeak important structural or functional features. Our novel methodol...
Topics
- AIDS Vaccines
- Amino Acid Sequence
- Computational Biology
- Computer Simulation
- Conserved Sequence
- Databases, Factual
- Drug Design
- Gene Products, env
- HIV
- Information Theory
- Markov Chains
- Mutation
- Protein Conformation
- Proteins
- Sequence Alignment
- Sequence Homology, Amino Acid
- Software
