LQM: Linguistically Motivated Multidimensional Quality Metrics for Machine Translation

Interactive error explorer · Findings of ACL 2026
LQM overview connecting six linguistic levels to translation error examples: sociolinguistics and code or register selection in Moroccan Arabic; pragmatics and illocutionary force in German; semantics and polysemy in French; morphosyntax and verbal features in Mandarin; orthography and typos in Korean; and graphetics and character encoding in Spanish.
LQM at a glance. Six linguistic levels, illustrated with translation errors across languages. SRC: source; TRG: reference translation; MT: machine translation.
LQM Error Explorer

About the project

LQM is a language-agnostic framework for diagnosing machine translation errors across six linguistic levels: sociolinguistics, pragmatics, semantics, morphosyntax, orthography, and graphetics. It captures dialect, cultural, and contextual issues that broad evaluation schemes can miss.

The project evaluates six large language models on a bidirectional corpus of 3,850 sentences spanning seven Arabic dialects, using expert span-level annotations and severity-weighted quality scores. This companion explorer lets you filter the annotations using the LQM or MQM taxonomy and inspect highlighted error spans.

In this explorer, we present LQM framework applied on the following languages/dialects: Egyptian Arabic Dialect (EGY), English (ENG), Jordanian Arabic Dialect (JOR), Mauritanian Arabic Dialect (MAU), Moroccan Arabic Dialect (MOR), Palestinian Arabic Dialect (PAL), Emerati Arabic Dialect (UAE), and Yemeni Arabic Dialect (YEM).

—
matching examples
—
matching errors
—
models represented
—
avg. errors / example

Filtered overview

Select a bar label to filter.

Examples by model

Errors by category

Errors by severity

Examples

Example
100%