Full metadata record
DC Field | Value | Language |
---|---|---|
dc.contributor.author | Suri, M. | en_US |
dc.contributor.author | Rini, S. | en_US |
dc.date.accessioned | 2019-09-02T07:45:42Z | - |
dc.date.available | 2019-09-02T07:45:42Z | - |
dc.date.issued | 2019-01-01 | en_US |
dc.identifier.isbn | 978-1-7281-0584-0 | en_US |
dc.identifier.issn | 2374-3212 | en_US |
dc.identifier.uri | http://hdl.handle.net/11536/152568 | - |
dc.description.abstract | In the Dictionary-based String Matching (DSM) problem, an Information Retrieval (IR) system has access to a source sequence and stores the position of a certain number of strings in a posting table. When a user inquires the position of a string, the IR system, instead of searching in the source sequence directly, relies on the the posting table to answer the query more efficiently. In this paper, the Statistical DSM problem is proposed as a statistical and information-theoretic formulation of the classic DSM problem in which both the source and the query have a statistical description while the strings stored in the posting sequence are described as a code. Through this formulation, we define the communication efficiency of the IR system as the average cost in retrieving the entries of the posting list from the posting table, in the limit of an infinitely long source sequence. This formulation is used to study the communication efficiency for the case in which the dictionary is composed of (i) all the strings of a given length, referred to as k-grams , and (ii) run-length codes. | en_US |
dc.language.iso | en_US | en_US |
dc.subject | Dictionary-based string matching | en_US |
dc.subject | Content based retrieval | en_US |
dc.subject | Indexing database | en_US |
dc.subject | Information retrieval | en_US |
dc.subject | Phrase searching | en_US |
dc.title | THE STATISTICAL DICTIONARY-BASED STRING MATCHING PROBLEM | en_US |
dc.type | Proceedings Paper | en_US |
dc.identifier.journal | IRAN WORKSHOP ON COMMUNICATION AND INFORMATION THEORY (IWCIT 2019) | en_US |
dc.citation.spage | 0 | en_US |
dc.citation.epage | 0 | en_US |
dc.contributor.department | 電機工程學系 | zh_TW |
dc.contributor.department | Department of Electrical and Computer Engineering | en_US |
dc.identifier.wosnumber | WOS:000476947400006 | en_US |
dc.citation.woscount | 0 | en_US |
Appears in Collections: | Conferences Paper |