The Repeat Pattern Toolkit (RPT): analyzing the structure and evolution of the C. elegans genome - PubMed (original) (raw)

Affiliations

The Repeat Pattern Toolkit (RPT): analyzing the structure and evolution of the C. elegans genome

P Agarwal et al. Proc Int Conf Intell Syst Mol Biol. 1994.

Abstract

Over 3.6 million bases of DNA sequence from chromosome III of the C. elegans have been determined. The availability of this extended region of contiguous sequence has allowed us to analyze the nature and prevalence of repetitive sequences in the genome of a eukaryotic organism with a high gene density. We have assembled a Repeat Pattern Toolkit (RPT) to analyze the patterns of repeats occurring in DNA. The tools include identifying significant local alignments (utilizing both two-way and three-way alignments), dividing the set of alignments into connected components (signifying repeat families), computing evolutionary distance between repeat family members, constructing minimum spanning trees from the connected components, and visualizing the evolution of the repeat families. Over 7000 families of repetitive sequences were identified. The size of the families ranged from isolated pairs to over 1600 segments of similar sequence. Approximately 12.3% of the analyzed sequence participates in a repeat element.

PubMed Disclaimer

Similar articles

Cited by

MeSH terms