DE eng

Search in the Catalogues and Directories

Hits 1 – 5 of 5

1
CSA++: Fast Pattern Search for Large Alphabets ...
Abstract: Indexed pattern search in text has been studied for many decades. For small alphabets, the FM-Index provides unmatched performance, in terms of both space required and search speed. For large alphabets -- for example, when the tokens are words -- the situation is more complex, and FM-Index representations are compact, but potentially slow. In this paper we apply recent innovations from the field of inverted indexing and document retrieval to compressed pattern search, including for alphabets into the millions. Commencing with the practical compressed suffix array structure developed by Sadakane, we show that the Elias-Fano code-based approach to document indexing can be adapted to provide new tradeoff options in indexed pattern search, and offers significantly faster pattern processing compared to previous implementations, as well as reduced space requirements. We report a detailed experimental evaluation that demonstrates the relative advantages of the new approach, using the standard Pizza&Chili ...
Keyword: Data Structures and Algorithms cs.DS; FOS Computer and information sciences
URL: https://arxiv.org/abs/1605.05404
https://dx.doi.org/10.48550/arxiv.1605.05404
BASE
Hide details
2
Frontiers, Challenges, and Opportunities for Information Retrieval – Report from SWIRL 2012, The Second Strategic Workshop on Information Retrieval in Lorne
Kelly, Diane; Clarke, Charles L.A.; Moffat, Alistair. - : KTH, Teoretisk datalogi, TCS, 2012. : ACM, 2012
BASE
Show details
3
Managing gigabytes : compressing and indexing documents and images
Witten, Ian H.; Moffat, Alistair; Bell, Thimothy C.. - San Francisco [etc.] : Morgan Kaufmann, 1999
MPI für Psycholinguistik
Show details
4
In situ generation of compressed inverted files
In: American Society for Information Science. Journal of the American Society for Information Science. - New York, NY : Wiley 46 (1995) 7, 537-550
BLLDB
Show details
5
Managing gigabytes : compressing and indexing documents and images
Moffat, Alistair; Witten, Ian H.. - New York : Van Nostrand Reinhold [u.a.], 1994
IDS Mannheim
Show details

Catalogues
0
1
0
0
0
0
0
Bibliographies
1
0
0
0
0
0
0
0
1
Linked Open Data catalogues
0
Online resources
0
0
0
0
Open access documents
2
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern