• Cite Icon1
  • https://doi.org/10.5220/0001935200050012Copy DOI Icon

English

  • Jan 1, 2008
  • Artur J Ferreira +2 more
Show More
  • Abstract
  • Literature Map
  • Citations
  • Similar Papers
Abstract

Keywords: Lempel-Ziv, Lossless Data Compression, Suffix Arrays, Suffix Tre es, String Matching.Abstract: Lossless compression algorithms of the Lempel-Ziv (LZ) family are widely used in a variety of applications.The LZ encoder and decoder exhibit a high asymmetry, regarding time and memory requirements, with theformer being much more demanding. Several techniques have been used to speed up the encoding process;among them is the use of suffix trees. In this paper, we explore the use of a simple data structure, namedsuffix array , to hold the dictionary of the LZ encoder, and propose an algorithm to search the dictionary.A comparison with the suffix tree based LZ encoder is carried out, showin g that the compression ratios areroughly the same. The ammount of memory required by the suffix arra y is fixed, being much lower than thevariable memory requirements of the suffix tree encoder, which depen ds on the text to encode. We concludethat suffix arrays are a very interesting option regarding the tradeoff b etween time, memory, and compressionratio, when compared with suffix trees, that make them preferable in som e compression scenarios.

Similar Papers
  • Supplementary Content
  • Citations1

Inverse Suffix Array Queries for 2-Dimensional Pattern Matching in Near-Compact Space

  • Jan 01, 2021
  • DROPS (Schloss Dagstuhl – Leibniz Center for Informatics)
  • Dhrumil Patel +1
  • Research Article
  • Citations81

Breaking a Time-and-Space Barrier in Constructing Full-Text Indices

  • Jan 01, 2009
  • SIAM Journal on Computing
  • Wing-Kai Hon +2
  • PDF
  • Research Article
  • Citations17

Suffix-Sorting via Shannon-Fano-Elias Codes

  • Apr 01, 2010
  • Algorithms
  • Donald Adjeroh +1
  • Conference Article
  • Citations42

Parallel suffix array and least common prefix for the GPU

  • Feb 23, 2013
  • Mrinal Deo +1
  • Book Chapter
  • Citations7

In-Place Sparse Suffix Sorting

  • Jan 01, 2018
  • Nicola Prezza
  • Conference Article
  • Citations58

String Matching in Hardware Using the FM-Index

  • May 01, 2011
  • Edward Fernandez +2
  • Research Article
  • Citations337

The string B-tree

  • Mar 01, 1999
  • Journal of the ACM
  • Paolo Ferragina +1
  • Conference Article
  • Citations1

An ACGT-Words Tree for Efficient Data Access in Genomic Databases

  • Apr 01, 2007
  • Ye-In Chang +3
  • Research Article
  • Citations25

The generalised k-Truncated Suffix Tree for time-and space-efficient searches in multiple DNA or protein sequences

  • Jan 01, 2008
  • International Journal of Bioinformatics Research and Applications
  • Marcel H. Schulz +2
  • Book Chapter
  • Citations1

Parallel Construction of Succinct Representations of Suffix Tree Topologies

  • Jan 01, 2015
  • Uwe Baier +2
  • Conference Article
  • Citations5

Sequence learning using the adaptive suffix trie algorithm

  • Jun 01, 2012
  • Upuli Gunasinghe +1
  • Research Article
  • Citations26

Efficient computation of shortest absent words in a genomic sequence

  • May 13, 2010
  • Information Processing Letters
  • Zong-Da Wu +2
  • Conference Article
  • Citations14

I/O-Efficient Compressed Text Indexes: From Theory to Practice

  • Jan 01, 2010
  • Sheng-Yuan Chiu +3
  • Research Article
  • Citations32

Efficient Maximal Repeat Finding Using the Burrows-Wheeler Transform and Wavelet Tree

  • Sep 27, 2011
  • IEEE/ACM Transactions on Computational Biology and Bioinformatics
  • M Oguzhan Kulekci +2
  • Research Article
  • Citations5

Error Tree: A Tree Structure for Hamming and Edit Distances and Wildcards Matching

  • Sep 24, 2015
  • Journal of Computational Biology
  • Anas Al-Okaily
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.