Improved space-time tradeoffs for approximate full-text indexing with one edit error

Belazzougui, Djamal

Computer Science > Data Structures and Algorithms

arXiv:1103.2167 (cs)

[Submitted on 10 Mar 2011 (v1), last revised 21 Aug 2014 (this version, v3)]

Title:Improved space-time tradeoffs for approximate full-text indexing with one edit error

Authors:Djamal Belazzougui

View PDF

Abstract:In this paper we are interested in indexing texts for substring matching queries with one edit error. That is, given a text $T$ of $n$ characters over an alphabet of size $\sigma$, we are asked to build a data structure that answers the following query: find all the $occ$ substrings of the text that are at edit distance at most $1$ from a given string $q$ of length $m$. In this paper we show two new results for this problem. The first result, suitable for an unbounded alphabet, uses $O(n\log^\epsilon n)$ (where $\epsilon$ is any constant such that $0<\epsilon<1$) words of space and answers to queries in time $O(m+occ)$. This improves simultaneously in space and time over the result of Cole et al. The second result, suitable only for a constant alphabet, relies on compressed text indices and comes in two variants: the first variant uses $O(n\log^{\epsilon} n)$ bits of space and answers to queries in time $O(m+occ)$, while the second variant uses $O(n\log\log n)$ bits of space and answers to queries in time $O((m+occ)\log\log n)$. This second result improves on the previously best results for constant alphabets achieved in Lam et al. (Algorithmica 2008) and Chan et al. (Algorithmica 2010).

Comments:	Accepted for publication in a journal (28 pages)
Subjects:	Data Structures and Algorithms (cs.DS)
Cite as:	arXiv:1103.2167 [cs.DS]
	(or arXiv:1103.2167v3 [cs.DS] for this version)
	https://doi.org/10.48550/arXiv.1103.2167

Submission history

From: Djamal Belazzougui [view email]
[v1] Thu, 10 Mar 2011 23:25:45 UTC (20 KB)
[v2] Thu, 17 Oct 2013 00:56:34 UTC (46 KB)
[v3] Thu, 21 Aug 2014 21:28:27 UTC (46 KB)

Computer Science > Data Structures and Algorithms

Title:Improved space-time tradeoffs for approximate full-text indexing with one edit error

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Data Structures and Algorithms

Title:Improved space-time tradeoffs for approximate full-text indexing with one edit error

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators