1998Unpublished venueOpen access

An algorithm for finding tandem repeats of unspecified pattern size

Gary Benson

Open full text 26 citations

Abstract

A tandem repeat is two or more contiguous, approtimate copies of a pattern of nucleotides.Tandem repeats occur frequently in the human genome.They have been shown to cause human disease, may play a variety of regulatory and evolutionary roles, and are important laboratory tools.Extensive knowledge about pattern sizes, copy number, mutational history, etc. for t.andem repeats has been limited because of the difficulty of detecting them in genomic sequence data.In this paper, me present a new algorithm for finding tandem repeats in DNA sequences without the need to specify either the pattern or pattern size.The algorithm is based on the detection of k-tuple matches.It uses a probabiitic model of tandem repeats and a collection of statistical criteria based on that modeL We demonstrate the algorithm's speed and its abiity to detect tandem repeats that have undergone extensive mutational change by analyzing 4 sequences in the 2OOKb to 700Kb range.

Open-access reader

About this research paper

What this paper is about

A tandem repeat is two or more contiguous, approtimate copies of a pattern of nucleotides.Tandem repeats occur frequently in the human genome.They have been shown to cause human disease, may play a variety of regulatory and evolutionary roles, and are important laboratory tools.Extensive knowledge about pattern sizes, copy number, mutational history, etc. for t.andem repeats has been limited because of the difficulty of detecting them in genomic sequence data.In this paper, me present a new algorithm for finding tandem repeats in DNA sequences without the need to specify either the pattern or pattern size.The algorithm is based on the detection of k-tuple matches.It uses a probabiitic model of tandem repeats and a collection of statistical criteria based on that modeL We demonstrate the algorithm's speed and its abiity to detect tandem repeats that have undergone extensive mutational change by analyzing 4 sequences in the 2OOKb to 700Kb range.

Why it matters

OpenAlex reports 26 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

A tandem repeat is two or more contiguous, approtimate copies of a pattern of nucleotides.Tandem repeats occur frequently in the human genome.They have been shown to cause human disease, may play a variety of regulatory and evolutionary roles, and are important laboratory tools.Extensive knowledge about pattern sizes, copy number, mutational history, etc. for t.andem repeats has been limited because of the difficulty of detecting them in genomic sequence data.In this paper, me present a new algorithm for finding tandem repeats in DNA sequences without the need to specify either the pattern or pattern size.The algorithm is based on the detection of k-tuple matches.It uses a probabiitic model of tandem repeats and a collection of statistical criteria based on that modeL We demonstrate the algorithm's speed and its abiity to detect tandem repeats that have undergone extensive mutational change by analyzing 4 sequences in the 2OOKb to 700Kb range.

Key concepts: Computer science, Algorithm, Tandem repeat, Tandem, Genetics, Engineering, Biology, Genome

Related papers

Back to paper searchBrowse research topicsOriginal source
An algorithm for finding tandem repeats of unspecified pattern size — Research Paper | ScholarLens