Summary of the paper

Title Manually Annotated Corpus of Polish Texts Published between 1830 and 1918
Authors Witold Kieraś and Marcin Woliński
Abstract The paper presents a manually annotated 625,000 tokens large historical corpus of -- fiction, drama, popular science, essays and newspapers of the period. The corpus provides three layers: transliteration, transcription and morphosyntactic annotation. The annotation process as well as the corpus itself are described in detail in the paper.
Topics Morphology, Corpus (Creation, Annotation, Etc.), Other
Full paper Manually Annotated Corpus of Polish Texts Published between 1830 and 1918
Bibtex @InProceedings{KIERAŚ18.675,
  author = {Witold Kieraś and Marcin Woliński},
  title = "{Manually Annotated Corpus of Polish Texts Published between 1830 and 1918}",
  booktitle = {Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018)},
  year = {2018},
  month = {May 7-12, 2018},
  address = {Miyazaki, Japan},
  editor = {Nicoletta Calzolari (Conference chair) and Khalid Choukri and Christopher Cieri and Thierry Declerck and Sara Goggi and Koiti Hasida and Hitoshi Isahara and Bente Maegaard and Joseph Mariani and Hélène Mazo and Asuncion Moreno and Jan Odijk and Stelios Piperidis and Takenobu Tokunaga},
  publisher = {European Language Resources Association (ELRA)},
  isbn = {979-10-95546-00-9},
  language = {english}
  }
Powered by ELDA © 2018 ELDA/ELRA