Skip to main navigation Skip to search Skip to main content

Lithuanian-latvian-lithuanian parallel corpus

  • Andrius Utka*
  • , Kristine Levane-Petrova
  • , Agne Bielinskiene
  • , Jolanta Kovalevskaite
  • , Erika Rimkute
  • , Daira Vevere
  • *Corresponding author for this work
  • Vytautas Magnus University
  • University of Latvia

Research output: Chapter in Book/Report/Conference proceedingConference paperResearchpeer-review

8 Citations (Scopus)

Abstract

The goal of the paper is to present different problems related to the building of Parallel Corpus for two small languages, namely, Latvian and Lithuanian. The Lithuanian-Latvian-Lithuania Parallel Corpus (LILA) will contain 8 million running words; will be bidirectional, aligned on the sentence level. The problems include identifying, acquiring, preparing, and aligning parallel texts.

Original languageEnglish
Title of host publicationHuman Language Technologies - The Baltic Perspective. Proceedings of the Fifth International Conference Baltic HLT 2012
PublisherIOS Press BV
Pages260-264
Number of pages5
ISBN (Print)9781614991328
DOIs
Publication statusPublished - 2012
Externally publishedYes
Event5th International Conference on Human Language Technologies - The Baltic Perspective, Baltic HLT 2012 - Tartu, Estonia
Duration: 4 Oct 20125 Oct 2012

Publication series

NameFrontiers in Artificial Intelligence and Applications
Volume247
ISSN (Print)0922-6389
ISSN (Electronic)1879-8314

Conference

Conference5th International Conference on Human Language Technologies - The Baltic Perspective, Baltic HLT 2012
Country/TerritoryEstonia
CityTartu
Period4/10/125/10/12

Keywords

  • Latvian
  • Lithuanian
  • parallel corpus

Fingerprint

Dive into the research topics of 'Lithuanian-latvian-lithuanian parallel corpus'. Together they form a unique fingerprint.

Cite this