Summary of the paper

Title Ontology-based Interoperation of Linguistic Tools for an Improved Lemma Annotation in Spanish
Authors Antonio Pareja-Lora and Guadalupe Aguado de Cea
Abstract In this paper, we present an ontology-based methodology and architecture for the comparison, assessment, combination (and, to some extent, also contrastive evaluation) of the results of different linguistic tools. More specifically, we describe an experiment aiming at the improvement of the correctness of lemma tagging for Spanish. This improvement was achieved by means of the standardisation and combination of the results of three different linguistic annotation tools (Bitext’s DataLexica, Connexor’s FDG Parser and LACELL’s POS tagger), using (1) ontologies, (2) a set of lemma tagging correction rules, determined empirically during the experiment, and (3) W3C standard languages, such as XML, RDF(S) and OWL. As we show in the results of the experiment, the interoperation of these tools by means of ontologies and the correction rules applied in the experiment improved significantly the quality of the resulting lemma tagging (when compared to the separate lemma tagging performed by each of the tools that we made interoperate).
Topics Corpus (creation, annotation, etc.), Evaluation methodologies, LR Infrastructures and Architectures
Full paper Ontology-based Interoperation of Linguistic Tools for an Improved Lemma Annotation in Spanish
Slides -
Bibtex @InProceedings{PAREJALORA10.92,
  author = {Antonio Pareja-Lora and Guadalupe Aguado de Cea},
  title = {Ontology-based Interoperation of Linguistic Tools for an Improved Lemma Annotation in Spanish},
  booktitle = {Proceedings of the Seventh International Conference on Language Resources and Evaluation (LREC'10)},
  year = {2010},
  month = {may},
  date = {19-21},
  address = {Valletta, Malta},
  editor = {Nicoletta Calzolari (Conference Chair) and Khalid Choukri and Bente Maegaard and Joseph Mariani and Jan Odijk and Stelios Piperidis and Mike Rosner and Daniel Tapias},
  publisher = {European Language Resources Association (ELRA)},
  isbn = {2-9517408-6-7},
  language = {english}
 }
Powered by ELDA © 2010 ELDA/ELRA