Summary of the paper

Title Second HAREM: Advancing the State of the Art of Named Entity Recognition in Portuguese
Authors Cláudia Freitas, Cristina Mota, Diana Santos, Hugo Gonçalo Oliveira and Paula Carvalho
Abstract In this paper, we present Second HAREM, the second edition of an evaluation campaign for Portuguese, addressing named entity recognition (NER). This second edition also included two new tracks: the recognition and normalization of temporal entities (proposed by a group of participants, and hence not covered on this paper) and ReRelEM, the detection of semantic relations between named entities. We summarize the setup of Second HAREM by showing the preserved distinctive features and discussing the changes compared to the first edition. Furthermore, we present the main results achieved and describe the available resources and tools developed under this evaluation, namely,(i) the golden collections, i.e. a set of documents whose named entities and semantic relations between those entities were manually annotated, (ii) the Second HAREM collection (which contains the unannotated version of the golden collection), as well as the participating systems results on it, (iii) the scoring tools, and (iv) SAHARA, a Web application that allows interactive evaluation. We end the paper by offering some remarks about what was learned.
Topics Named Entity recognition, Information Extraction, Information Retrieval, Evaluation methodologies
Full paper Second HAREM: Advancing the State of the Art of Named Entity Recognition in Portuguese
Slides Second HAREM: Advancing the State of the Art of Named Entity Recognition in Portuguese
Bibtex @InProceedings{FREITAS10.412,
  author = {Cláudia Freitas and Cristina Mota and Diana Santos and Hugo Gonçalo Oliveira and Paula Carvalho},
  title = {Second HAREM: Advancing the State of the Art of Named Entity Recognition in Portuguese},
  booktitle = {Proceedings of the Seventh International Conference on Language Resources and Evaluation (LREC'10)},
  year = {2010},
  month = {may},
  date = {19-21},
  address = {Valletta, Malta},
  editor = {Nicoletta Calzolari (Conference Chair) and Khalid Choukri and Bente Maegaard and Joseph Mariani and Jan Odijk and Stelios Piperidis and Mike Rosner and Daniel Tapias},
  publisher = {European Language Resources Association (ELRA)},
  isbn = {2-9517408-6-7},
  language = {english}
 }
Powered by ELDA © 2010 ELDA/ELRA