Papers
arxiv:1910.07468

Fine-grained evaluation of Quality Estimation for Machine translation based on a linguistically-motivated Test Suite

Published on Oct 16, 2019
Authors:
,
,
,

Abstract

A linguistically motivated test suite with categorized translation errors evaluates quality estimation systems by measuring their ability to distinguish correct from erroneous outputs.

We present an alternative method of evaluating Quality Estimation systems, which is based on a linguistically-motivated Test Suite. We create a test-set consisting of 14 linguistic error categories and we gather for each of them a set of samples with both correct and erroneous translations. Then, we measure the performance of 5 Quality Estimation systems by checking their ability to distinguish between the correct and the erroneous translations. The detailed results are much more informative about the ability of each system. The fact that different Quality Estimation systems perform differently at various phenomena confirms the usefulness of the Test Suite.

Community

Sign up or log in to comment

Get this paper in your agent:

hf papers read 1910.07468
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper 1

Datasets citing this paper 0

No dataset linking this paper

Cite arxiv.org/abs/1910.07468 in a dataset README.md to link it from this page.

Spaces citing this paper 0

No Space linking this paper

Cite arxiv.org/abs/1910.07468 in a Space README.md to link it from this page.

Collections including this paper 0

No Collection including this paper

Add this paper to a collection to link it from this page.