Title |
Named and Specific Entity Detection in Varied Data: The Quæro Named Entity Baseline Evaluation |
Authors |
Olivier Galibert, Ludovic Quintard, Sophie Rosset, Pierre Zweigenbaum, Claire Nédellec, Sophie Aubin, Laurent Gillard, Jean-Pierre Raysz, Delphine Pois, Xavier Tannier, Louise Deléger and Dominique Laurent |
Abstract |
The Quæro program that promotes research and industrial innovation on technologies for automatic analysis and classification of multimedia and multilingual documents. Within its context a set of evaluations of Named Entity recognition systems was held in 2009. Four tasks were defined. The first two concerned traditional named entities in French broadcast news for one (a rerun of ESTER 2) and of OCR-ed old newspapers for the other. The third was a gene and protein name extraction in medical abstracts. The last one was the detection of references in patents. Four different partners participated, giving a total of 16 systems. We provide a synthetic descriptions of all of them classifying them by the main approaches chosen (resource-based, rules-based or statistical), without forgetting the fact that any modern system is at some point hybrid. The metric (the relatively standard Slot Error Rate) and the results are also presented and discussed. Finally, a process is ongoing with preliminary acceptance of the partners to ensure the availability for the community of all the corpora used with the exception of the non-Quæro produced ESTER 2 one. |
Topics |
Named Entity recognition |
Full paper |
Named and Specific Entity Detection in Varied Data: The Quæro Named Entity Baseline Evaluation |
Slides |
- |
Bibtex |
@InProceedings{GALIBERT10.191,
author = {Olivier Galibert and Ludovic Quintard and Sophie Rosset and Pierre Zweigenbaum and Claire Nédellec and Sophie Aubin and Laurent Gillard and Jean-Pierre Raysz and Delphine Pois and Xavier Tannier and Louise Deléger and Dominique Laurent}, title = {Named and Specific Entity Detection in Varied Data: The Quæro Named Entity Baseline Evaluation}, booktitle = {Proceedings of the Seventh International Conference on Language Resources and Evaluation (LREC'10)}, year = {2010}, month = {may}, date = {19-21}, address = {Valletta, Malta}, editor = {Nicoletta Calzolari (Conference Chair) and Khalid Choukri and Bente Maegaard and Joseph Mariani and Jan Odijk and Stelios Piperidis and Mike Rosner and Daniel Tapias}, publisher = {European Language Resources Association (ELRA)}, isbn = {2-9517408-6-7}, language = {english} } |