Structured local exponential models for machine translation

dc.contributor.advisorResnik, Philipen_US
dc.contributor.authorSubotin, Michaelen_US
dc.contributor.departmentLinguisticsen_US
dc.contributor.publisherDigital Repository at the University of Marylanden_US
dc.contributor.publisherUniversity of Maryland (College Park, Md.)en_US
dc.date.accessioned2011-12-01T06:30:35Z
dc.date.available2011-12-01T06:30:35Z
dc.date.issued2010en_US
dc.description.abstractThis thesis proposes a synthesis and generalization of local exponential translation models, the subclass of feature-rich translation models which associate probability distributions with individual rewrite rules used by the translation system, such as synchronous context-free rules, or with other individual aspects of translation hypotheses such as word pairs or reordering events. Unlike other authors we use these estimates to replace the traditional phrase models and lexical scores, rather than in addition to them, thereby demonstrating that the local exponential phrase models can be regarded as a generalization of standard methods not only in theoretical but also in practical terms. We further introduce a form of local translation models that combine features associated with surface forms of rules and features associated with less specific representation -- including those based on lemmas, inflections, and reordering patterns -- such that surface-form estimates are recovered as a special case of the model. Crucially, the proposed approach allows estimation of parameters for the latter type of features from training sets that include multiple source phrases, thereby overcoming an important training set fragmentation problem which hampers previously proposed local translation models. These proposals are experimentally validated. Conditioning all phrase-based probabilities in a hierarchical phrase-based system on source-side contextual information produces significant performance improvements. Extending the contextually-sensitive estimates with features modeling source-side morphology and reordering patterns yields consistent additional improvements, while further experiments show significant improvements obtained from modeling observed and unobserved inflections for a morphologically rich target language.en_US
dc.identifier.urihttp://hdl.handle.net/1903/12150
dc.subject.pqcontrolledComputer scienceen_US
dc.subject.pqcontrolledLinguisticsen_US
dc.subject.pqcontrolledArtificial intelligenceen_US
dc.subject.pquncontrolledExponential modelsen_US
dc.subject.pquncontrolledMachine learningen_US
dc.subject.pquncontrolledMachine translationen_US
dc.subject.pquncontrolledMaximum entropyen_US
dc.titleStructured local exponential models for machine translationen_US
dc.typeThesisen_US

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
Subotin_umd_0117N_11851.pdf
Size:
321.7 KB
Format:
Adobe Portable Document Format