Predictor-estimator: Neuralquality estimation based on target word prediction for machine translation

Kim, H.; Jung, H.-Y.; Kwon, H.; Lee, J.-H.; Na, Seung-Hoon

doi:10.1145/3109480

Scholarworks@UNIST

UNIST Library

File Download

There are no files associated with this item.

SFX Link

Find it @ UNIST can give you direct access to the published full text of this article. (UNISTARs only)

Related Researcher

나승훈

Na, Seung-Hoon: Natural Language Processing Lab

Read More

Views & Downloads

Detailed Information

Cited time in webofscience

Cited time in scopus

Metadata Downloads

Predictor-estimator: Neuralquality estimation based on target word prediction for machine translation

Author(s): Kim, H., Jung, H.-Y., Kwon, H., Lee, J.-H., Na, Seung-Hoon

Issued Date: 2017-03

DOI: 10.1145/3109480

URI: https://scholarworks.unist.ac.kr/handle/201301/86820

Fulltext: https://dl.acm.org/doi/10.1145/3109480

Citation: ACM TRANSACTIONS ON ASIAN AND LOW-RESOURCE LANGUAGE INFORMATION PROCESSING, v.17, no.1, pp.3

Abstract: Recently, quality estimation has been attracting increasing interest from machine translation researchers, aiming at finding a good estimator for the quality of machine translation output. The common approach for quality estimation is to treat the problem as a supervised regression/classification task using a qualityannotated noisy parallel corpus, called quality estimation data, as training data. However, the available size of quality estimation data remains small, due to the too-expensive cost of creating such data. In addition, most conventional quality estimation approaches rely on manually designed features to model nonlinear relationships between feature vectors and corresponding quality labels. To overcome these problems, this article proposes a novel neural network architecture for quality estimation task-called the predictor-estimator-that considersword prediction as an additional pre-task. The major component of the proposed neural architecture is a word prediction model based on a modified neural machine translation model-a probabilistic model for predicting a targetword conditioned on all the other source and target contexts. The underlying assumption is that the word prediction model is highly related to quality estimation models and is therefore able to transfer useful knowledge to quality estimation tasks. Our proposed quality estimation method sequentially trains the following two types of neural models: (1) Predictor: a neural word prediction model trained from parallel corpora and (2) Estimator: a neural quality estimation model trained fromquality estimation data. To transferword a prediction task to a quality estimation task, we generate quality estimation feature vectors from theword prediction model and feed them into the quality estimation model. The experimental results on WMT15 and 16 quality estimation datasets show that our proposed method has great potential in the various sub-challenges. © 2017 ACM.

Publisher: Association for Computing Machinery

ISSN: 2375-4699

Keyword (Author): Feature extraction, Machine translation, Neural networks, Quality estimation, Word prediction, Bidirectional language model

Show Full Item Record

qrcode

RSS 1.0 RSS 2.0

UNIST | Library

Tel : 052-217-1403 / Email : scholarworks@unist.ac.kr

ScholarWorks@UNIST was established as an OAK Project for the National Library of Korea.