the.ai

Multilingual / Methods

verified

Machine Translation

The task that produced the encoder-decoder, attention, and eventually the transformer. It is also the task where evaluation is hardest to take seriously: there are many correct translations of a sentence, and the standard automatic metric compares against one or two of them by counting shared word sequences.

Viz primitive · budget-splitreference-overlap = 6

reference-overlap holds 25% of the budget; rest holds the remaining 75%.

Output that matches the reference's exact phrasing against output that conveys the same meaning differently, in n-grams. Drag the overlap up to watch the score rise — the second kind of output is correct and scores nothing.

6

Reviewed by opendroid · 2026-08-18

  • arXiv:2207.04672 — No Language Left Behind: Scaling Human-Centered Machine Translation