PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
July 19, 2002Genome Research1,054 citationsOpen Access

Automated De Novo Identification of Repeat Sequence Families in Sequenced Genomes

ZBZhirong BaoSESean R. Eddy

Key Points

Key points are not available for this paper at this time.

Abstract

Repetitive sequences make up a major part of eukaryotic genomes. We have developed an approach for the de novo identification and classification of repeat sequence families that is based on extensions to the usual approach of single linkage clustering of local pairwise alignments between genomic sequences. Our extensions use multiple alignment information to define the boundaries of individual copies of the repeats and to distinguish homologous but distinct repeat element families. When tested on the human genome, our approach was able to properly identify and group known transposable elements. The program, should be useful for first-pass automatic classification of repeats in newly sequenced genomes.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Bao et al. (2002) studied this question.

synapsesocial.com/papers/69d7faae66a29169b4bedacdhttps://doi.org/10.1101/gr.88502
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Amino acid substitution matrices from an information theoretic perspective1991 · 645 citations
  2. 2The algorithm design manual1998 · 914 citations
  3. 3Free left arms as precursor molecules in the evolution of Alu sequences1991 · 52 citations
  4. 4Transposable Elements and Genome Organization: A Comprehensive Survey of Retrotransposons Revealed by the Complete Saccharomyces cerevisiae Genome Sequence1998 · 548 citations
  5. 5DIALIGN 2: improvement of the segment-to-segment approach to multiple sequence alignment.1999 · 686 citations