Back to Search Start Over

Efficient accurate syntactic direct translation models: one tree at a time.

Authors :
Hassan, Hany
Sima'an, Khalil
Way, Andy
Source :
Machine Translation; Mar2012, Vol. 26 Issue 1/2, p121-136, 16p
Publication Year :
2012

Abstract

A challenging aspect of Statistical Machine Translation from Arabic to English lies in bringing the Arabic source morpho-syntax to bear on the lexical as well as word-order choices of the English target string. In this article, we extend the feature-rich discriminative Direct Translation Model 2 (DTM2) with a novel linear-time parsing algorithm based on an eager, incremental interpretation of Combinatory Categorial Grammar. This way we can reap the benefits of a target syntactic enhancement that leads to more grammatical output while also enabling dynamic decoding without the risk of blowing up decoding space and time requirements. Our model defines a mix of model parameters, some of which involve DTM2 source morpho-syntactic features, and others are novel target side syntactic features. Alongside translation features extracted from the derived parse tree, we explore syntactic features extracted from the incremental derivation process. Our empirical experiments show that our model significantly outperforms the state-of-the-art DTM2 system. [ABSTRACT FROM AUTHOR]

Details

Language :
English
ISSN :
09226567
Volume :
26
Issue :
1/2
Database :
Complementary Index
Journal :
Machine Translation
Publication Type :
Academic Journal
Accession number :
71861338
Full Text :
https://doi.org/10.1007/s10590-011-9116-7