George Wang
All publications
arXiv Preprint

MAATS: A Multi-Agent Automated Translation System Based on MQM Evaluation

Wang, X.; Hu, J.; Ali, S.
arXiv preprint, 2025
MAATS multi-agent translation system teaser
Abstract

We present MAATS, a Multi-Agent Automated Translation System that leverages the Multidimensional Quality Metrics (MQM) framework as a fine-grained signal for error detection and refinement. MAATS employs multiple specialized AI agents, each focused on a distinct MQM category (e.g., Accuracy, Fluency, Style, Terminology), followed by a synthesis agent that integrates the annotations to iteratively refine translations. This design contrasts with conventional single-agent methods that rely on self-correction.

Evaluated across diverse language pairs and Large Language Models (LLMs), MAATS outperforms zero-shot and single-agent baselines with statistically significant gains in both automatic metrics and human assessments. It excels particularly in semantic accuracy, locale adaptation, and linguistically distant language pairs. Qualitative analysis highlights its strengths in multi-layered error diagnosis, omission detection across perspectives, and context-aware refinement.

By aligning modular agent roles with interpretable MQM dimensions, MAATS narrows the gap between black-box LLMs and human translation workflows, shifting focus from surface fluency to deeper semantic and contextual fidelity.

Paper
Download PDF
Paper page 1 Paper page 2 Paper page 3 Paper page 4 Paper page 5 Paper page 6 Paper page 7 Paper page 8 Paper page 9 Paper page 10 Paper page 11 Paper page 12 Paper page 13 Paper page 14 Paper page 15 Paper page 16 Paper page 17 Paper page 18 Paper page 19 Paper page 20 Paper page 21 Paper page 22 Paper page 23 Paper page 24 Paper page 25 Paper page 26
Citation
@misc{wang2025maats, author = {Wang, George and Hu, Jiaqian and Ali, Safinah}, title = {MAATS: A Multi-Agent Automated Translation System Based on MQM Evaluation}, year = {2025}, eprint = {2505.14848}, archivePrefix = {arXiv} }