Title |
Automatic Summarization Using Terminological and Semantic Resources |
Authors |
Jorge Vivaldi, Iria da Cunha, Juan Manuel Torres-Moreno and Patricia Velázquez-Morales |
Abstract |
This paper presents a new algorithm for automatic summarization of specialized texts combining terminological and semantic resources: a term extractor and an ontology. The term extractor provides the list of the terms that are present in the text together their corresponding termhood. The ontology is used to calculate the semantic similarity among the terms found in the main body and those present in the document title. The general idea is to obtain a relevance score for each sentence taking into account both the termhood of the terms found in such sentence and the similarity among such terms and those terms present in the title of the document. The phrases with the highest score are chosen to take part of the final summary. We evaluate the algorithm with Rouge, comparing the resulting summaries with the summaries of other summarizers. The sentence selection algorithm was also tested as part of a standalone summarizer. In both cases it obtains quite good results although the perception is that there is a space for improvement. |
Topics |
Summarisation, Ontologies, Other |
Full paper |
Automatic Summarization Using Terminological and Semantic Resources |
Slides |
- |
Bibtex |
@InProceedings{VIVALDI10.400,
author = {Jorge Vivaldi and Iria da Cunha and Juan Manuel Torres-Moreno and Patricia Velázquez-Morales}, title = {Automatic Summarization Using Terminological and Semantic Resources}, booktitle = {Proceedings of the Seventh International Conference on Language Resources and Evaluation (LREC'10)}, year = {2010}, month = {may}, date = {19-21}, address = {Valletta, Malta}, editor = {Nicoletta Calzolari (Conference Chair) and Khalid Choukri and Bente Maegaard and Joseph Mariani and Jan Odijk and Stelios Piperidis and Mike Rosner and Daniel Tapias}, publisher = {European Language Resources Association (ELRA)}, isbn = {2-9517408-6-7}, language = {english} } |