• Home
  • Search
  • Computing all repeats using suffix arrays
  • Cite Icon29
  • https://doi.org/10.5555/998223.998227Copy DOI Icon

Computing all repeats using suffix arrays

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

We describe an algorithm that identifies all the repeating substrings (tandem, overlapping, and split) in a given string x = X[1..n]. Given the suffix arrays of x and of the reversed string x, the algorithm requires Θ(n) time for its execution and represents its output in Θ(n) space, either as a reduced suffix array (called an NE array) or as a reduced suffix tree (called an NE tree). The output substrings u are nonextendible (NE); that is, any extension of some occurrence of u in x, either to the left or to the right, yields a string (λu or uλ) that is unequal to the same extension of some other occurrence of u. Thus the number of substrings output is the minimum required to identify all the repeating substrings in x. The output can be used in a straightforward way to identify only repeating substrings that satisfy some proximity or minimum length condition.

Similar Papers
  • Supplementary Content
  • Citations1

Inverse Suffix Array Queries for 2-Dimensional Pattern Matching in Near-Compact Space

  • Jan 01, 2021
  • DROPS (Schloss Dagstuhl – Leibniz Center for Informatics)
  • Dhrumil Patel +1
  • PDF
  • Research Article
  • Citations17

Suffix-Sorting via Shannon-Fano-Elias Codes

  • Apr 01, 2010
  • Algorithms
  • Donald Adjeroh +1
  • Research Article
  • Citations81

Breaking a Time-and-Space Barrier in Constructing Full-Text Indices

  • Jan 01, 2009
  • SIAM Journal on Computing
  • Wing-Kai Hon +2
  • Conference Article
  • Citations42

Parallel suffix array and least common prefix for the GPU

  • Feb 23, 2013
  • Mrinal Deo +1
  • Conference Article
  • Citations1

English

  • Jan 01, 2008
  • Artur J Ferreira +2
  • Book Chapter
  • Citations7

In-Place Sparse Suffix Sorting

  • Jan 01, 2018
  • Nicola Prezza
  • Conference Article
  • Citations58

String Matching in Hardware Using the FM-Index

  • May 01, 2011
  • Edward Fernandez +2
  • Book Chapter
  • Citations26

Suffix Arrays on Words

  • Jul 09, 2007
  • Paolo Ferragina +1
  • Conference Article
  • Citations1

Search Results Clustering Based on Suffix Array and VSM

  • Dec 01, 2010
  • Shunlai Bai +3
  • Research Article
  • Citations30

Time-space trade-offs for compressed suffix arrays

  • Oct 31, 2001
  • Information Processing Letters
  • S.Srinivasa Rao
  • Research Article
  • Citations337

The string B-tree

  • Mar 01, 1999
  • Journal of the ACM
  • Paolo Ferragina +1
  • Conference Article
  • Citations1

An ACGT-Words Tree for Efficient Data Access in Genomic Databases

  • Apr 01, 2007
  • Ye-In Chang +3
  • Research Article
  • Citations25

The generalised k-Truncated Suffix Tree for time-and space-efficient searches in multiple DNA or protein sequences

  • Jan 01, 2008
  • International Journal of Bioinformatics Research and Applications
  • Marcel H. Schulz +2
  • Book Chapter
  • Citations1

Parallel Construction of Succinct Representations of Suffix Tree Topologies

  • Jan 01, 2015
  • Uwe Baier +2
  • Conference Article
  • Citations5

Sequence learning using the adaptive suffix trie algorithm

  • Jun 01, 2012
  • Upuli Gunasinghe +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.