首页 > AI前沿 > FACET at WMT 2026 Automated Translation Quality Evaluation Task

FACET at WMT 2026 Automated Translation Quality Evaluation Task

arXiv自然语言 2026-09-08 22:49 6 阅读 查看原文

Different error types in machine translation require different evidence.

Whether meaning is preserved can be judged only against the source, while whether the target is well-formed, or whether it names one entity consistently, can be judged from the target alone.

We present FACET

We present FACET, our reference-free submission to the WMT26 Automated Translation Quality Evaluation Task, which decomposes evaluation into Fluency, Accuracy, and Consistency passes and gives each pass only the context its error type requires.

A single fixed model is prompted three times, and the merged error spans yield the three task outputs, error spans, quality scores, and error-free labels, with no trained components.

We also submit FACET-C

We also submit FACET-C, which omits the Consistency pass.

Without gold labels, we characterize the predictions of FACET.

Its system rankings place post-edited human translation first, and the Consistency pass changes about a tenth of segment scores while leaving the ranking nearly unchanged.