Skip to main content
Springer Proceedings in Mathematics and StatisticsVolume 318, 2020, Pages 377-391International conference in honor of the 90th Birthday of Constantin Corduneanu, CONCORD-90 2018; Ekaterinburg; Russian Federation; 26 July 2018 through 31 July 2018; Code 240419

Proximity Full-Text Searches of Frequently Occurring Words with a Response Time Guarantee(Conference Paper)(Open Access)

  Save all to author list
  • aUral Federal University, Lenina 51, Yekaterinburg, 620083, Russian Federation
  • bINSM, Yekaterinburg, Russian Federation

Abstract

Full-text search engines are important tools for information retrieval. In a proximity full-text search, a document is relevant if it contains query terms near each other, especially if the query terms are frequently occurring words. For each word in the text, we use additional indexes to store information about nearby words at distances from the given word of less than or equal to MaxDistance, which is a parameter. A search algorithm for the case when the query consists of high-frequently occurring words is discussed. In addition, we present results of experiments with different values of MaxDistance to evaluate the search speed dependence on the value of MaxDistance. These results show that the average time of the query execution with our indexes is 94.7–45.9 times (depending on the value of MaxDistance) less than that with standard inverted files when queries that contain high-frequently occurring words are evaluated. © Springer Nature Switzerland AG 2020.

Author keywords

Additional indexesFull-text searchInformation retrievalInverted indexesProximity searchSearch enginesTerm proximity

Indexed keywords

Engineering uncontrolled termsFull-text searchFull-text search enginesInverted filesQuery executionQuery termsResponse-time guaranteesSearch AlgorithmsSearch speed
Engineering main heading:Search engines
  • ISSN: 21941009
  • ISBN: 978-303042175-5
  • Source Type: Conference Proceeding
  • Original language: English
  • DOI: 10.1007/978-3-030-42176-2_37
  • Document Type: Conference Paper
  • Volume Editors: Pinelas S.,Kim A.,Vlasov V.
  • Publisher: Springer

  Veretennikov, A.B.; INSM, Yekaterinburg, Russian Federation;
© Copyright 2020 Elsevier B.V., All rights reserved.

Cited by 0 documents

{"topic":{"name":"Information Retrieval; Query Processing; Caching","id":17537,"uri":"Topic/17537","prominencePercentile":79.66451,"prominencePercentileString":"79.665","overallScholarlyOutput":0},"dig":"70d9384c9966e9543afd577e48c857672b96fb155bb015d7cab35bf398036346"}

SciVal Topic Prominence

Topic:
Prominence percentile: