Skip to main navigation Skip to search Skip to main content

What can we learn from almost a decade of food tweets

  • Uga Sproagis*
  • , Matīss Rikters
  • *Corresponding author for this work
    • University of Latvia
    • The University of Tokyo
    • Datorikas fakultāte Maģistrantūra

    Research output: Chapter in Book/Report/Conference proceedingConference paperResearchpeer-review

    5 Citations (Scopus)

    Abstract

    We present the Latvian Twitter Eater Corpus - a set of tweets in the narrow domain related to food, drinks, eating and drinking. The corpus has been collected over time-span of over 8 years and includes over 2 million tweets entailed with additional useful data. We also separate two sub-corpora of question and answer tweets and sentiment annotated tweets. We analyse the contents of the corpus and demonstrate use-cases for the sub-corpora by training domain-specific question-answering and sentiment-analysis models using the data from the corpus.

    Original languageEnglish
    Title of host publicationHuman Language Technologies - The Baltic Perspective - Proceedings of the 9th International Conference Baltic HLT 2020
    EditorsAndrius Utka, Jurgita Vaicenoniene, Jolanta Kovalevskaite, Danguole Kalinauskaite
    Place of PublicationAmsterdam
    PublisherIOS Press
    Pages191-198
    ISBN (Print)9781643681160
    DOIs
    Publication statusPublished - 15 Sept 2020

    Publication series

    NameFrontiers in Artificial Intelligence and Applications
    Volume328
    ISSN (Print)0922-6389
    ISSN (Electronic)1879-8314

    OECD Field of Science

    • 1.2 Computer and Information Sciences

    Keywords

    • Annotated corpora
    • Food data
    • Latvian
    • Social networks

    Fingerprint

    Dive into the research topics of 'What can we learn from almost a decade of food tweets'. Together they form a unique fingerprint.

    Cite this