Skip to main navigation Skip to search Skip to main content

Text retrieval by term co-occurrences in a query-based vector space

  • Eriks Sneiders*
  • *Corresponding author for this work
  • Stockholm University

Research output: Chapter in Book/Report/Conference proceedingConference paperResearchpeer-review

4 Citations (Scopus)

Abstract

Term co-occurrence in a sentence or paragraph is a powerful and often overlooked feature for text matching in document retrieval. In our experiments with matching email-style query messages to webpages, such term co-occurrence helped greatly to filter and rank documents, compared to matching document-size bags-of-words. The paper presents the results of the experiments as well as a text-matching model where the query shapes the vector space, a document is modelled by two or three vectors in this vector space, and the query-document similarity score depends on the length of the vectors and the relationships between them.

Original languageEnglish
Title of host publicationCOLING 2016 - 26th International Conference on Computational Linguistics, Proceedings of COLING 2016
Subtitle of host publicationTechnical Papers
PublisherAssociation for Computational Linguistics, ACL Anthology
Pages2356-2365
Number of pages10
ISBN (Print)9784879747020
Publication statusPublished - 2016
Externally publishedYes
Event26th International Conference on Computational Linguistics, COLING 2016 - Osaka, Japan
Duration: 11 Dec 201616 Dec 2016

Publication series

NameCOLING 2016 - 26th International Conference on Computational Linguistics, Proceedings of COLING 2016: Technical Papers

Conference

Conference26th International Conference on Computational Linguistics, COLING 2016
Country/TerritoryJapan
CityOsaka
Period11/12/1616/12/16

Fingerprint

Dive into the research topics of 'Text retrieval by term co-occurrences in a query-based vector space'. Together they form a unique fingerprint.

Cite this