Title |
Alignment-based reordering for SMT |
Authors |
Maria Holmqvist, Sara Stymne, Lars Ahrenberg and Magnus Merkel |
Abstract |
We present a method for improving word alignment quality for phrase-based statistical machine translation by reordering the source text according to the target word order suggested by an initial word alignment. The reordered text is used to create a second word alignment which can be an improvement of the first alignment, since the word order is more similar. The method requires no other pre-processing such as part-of-speech tagging or parsing. We report improved Bleu scores for English-to-German and English-to-Swedish translation. We also examined the effect on word alignment quality and found that the reordering method increased recall while lowering precision, which partly can explain the improved Bleu scores. A manual evaluation of the translation output was also performed to understand what effect our reordering method has on the translation system. We found that where the system employing reordering differed from the baseline in terms of having more words, or a different word order, this generally led to an improvement in translation quality. |
Topics |
Multilinguality, Corpus (creation, annotation, etc.), Machine Translation, SpeechToSpeech Translation |
Full paper |
Alignment-based reordering for SMT |
Bibtex |
@InProceedings{HOLMQVIST12.1000,
author = {Maria Holmqvist and Sara Stymne and Lars Ahrenberg and Magnus Merkel}, title = {Alignment-based reordering for SMT}, booktitle = {Proceedings of the Eight International Conference on Language Resources and Evaluation (LREC'12)}, year = {2012}, month = {may}, date = {23-25}, address = {Istanbul, Turkey}, editor = {Nicoletta Calzolari (Conference Chair) and Khalid Choukri and Thierry Declerck and Mehmet Uğur Doğan and Bente Maegaard and Joseph Mariani and Asuncion Moreno and Jan Odijk and Stelios Piperidis}, publisher = {European Language Resources Association (ELRA)}, isbn = {978-2-9517408-7-7}, language = {english} } |