Genome Informatics

A novel computational method for the identification of candidate proteins useful as anti-infectives

(TG/CA)n Repeats in Human Housekeeping Genes

The unravelling of human genome sequence gives a new opportunity to investigate the role of repetitive sequences in gene regulation. Among the various types of repetitive sequences, the dinucleotide (TG:CA)(n) repeats are one of the most abundant in …

Nonrandom Distribution of Alu Elements in Genes of Various Functional Categories: Insight from Analysis of Human Chromosomes 21 and 22

The first draft of the human genome has revealed enormous variability in the global distribution of Alu repeat elements. There are regions such as the four homeobox gene clusters, which are nearly devoid of these repeats that contrast with repeat …

A Novel Complexity Measure for Comparative Analysis of Protein Sequences from Complete Genomes

Analysis of sequence complexities of proteins is an important step in the characterization and classification of new genomes. A new measure has been proposed to compute sequence complexity in protein sequences based on linguistic complexity. The …

Comparative genomics using data mining tools

We have analysed the genomes of representatives of three kingdoms of life, namely, archaea, eubacteria and eukaryota using data mining tools based on compositional analyses of the protein sequences. The representatives chosen in this analysis were …