Title |
Application of Resource-based Machine Translation to Real Business Scenes |
Authors |
Hitoshi Isahara, Masao Utiyama, Eiko Yamamoto, Akira Terada and Yasunori Abe |
Abstract |
As huge quantities of documents have become available, services using natural language processing technologies trained by huge corpora have emerged, such as information retrieval and information extraction. In this paper we verify the usefulness of resource-based, or corpus-based, translation in the aviation domain as a real business situation. This study is important from both a business perspective and an academic perspective. Intuitively, manuals for similar products, or manuals for different versions of the same product, are likely to resemble each other. Therefore, even with only a small training data, a corpus-based MT system can output useful translations. The corpus-based approach is powerful when the target is repetitive. Manuals for similar products, or manuals for different versions of the same product, are real-world documents that are repetitive. Our experiments on translation of manual documents are still in a beginning stage. However, the BLEU score from very small number of training sentences is already rather high. We believe corpus-based machine translation is a player full of promise in this kind of actual business scene. |
Language |
Multiple languages |
Topics |
Machine Translation, SpeechToSpeech Translation, Multilinguality, Usability, user satisfaction |
Full paper |
Application of Resource-based Machine Translation to Real Business Scenes |
Slides |
- |
Bibtex |
@InProceedings{ISAHARA08.780,
author = {Hitoshi Isahara, Masao Utiyama, Eiko Yamamoto, Akira Terada and Yasunori Abe},
title = {Application of Resource-based Machine Translation to Real Business Scenes},
booktitle = {Proceedings of the Sixth International Conference on Language Resources and Evaluation (LREC'08)},
year = {2008},
month = {may},
date = {28-30},
address = {Marrakech, Morocco},
editor = {Nicoletta Calzolari (Conference Chair), Khalid Choukri, Bente Maegaard, Joseph Mariani, Jan Odijk, Stelios Piperidis, Daniel Tapias},
publisher = {European Language Resources Association (ELRA)},
isbn = {2-9517408-4-0},
note = {http://www.lrec-conf.org/proceedings/lrec2008/},
language = {english}
} |