MT Evaluation

MT: Machine Translation

Latest papers 106

All topics
CardsList
  1. Generative AI translations in high-stakes emergency messaging

    Oct 6, 2026Nune Ayvazan, Anthony Pym, Yu HaoDisaster ResponseMachine Translation

  2. Precision over Scale: A Polish-Silesian Benchmark and a Translation System Outperforming Open-Source and Commercial Models

    Oct 1, 2026Grzegorz Kulik, Mikołaj Pokrywka, Adam Jatowt +1Low-Resource MTMachine Translation

  3. TermJudge: A Document-Level Metric Judging, Not Counting, Terminology in Machine Translation Evaluation

    Sep 28, 2026Nicolas Dahan, Fran{\cc}ois Yvon, Rachel BawdenAutomated EvaluationDocument-Level MT

  4. TTLab at AlexandriaX-2026: A Fine-Tuned Surface Tagger for Arabic Machine-Translation Error-Span Detection and Classification

    Sep 24, 2026Ali Abusaleh, Bhuvanesh Verma, Alexander MehlerArabic NLPMachine Translation

  5. COILD: An Indic-Centric Parallel Corpus and Benchmark for Machine Translation Across Indian Languages

    Sep 23, 2026Kshetrimayum Boynao Singh, Nitin Kumar Mishra, Palash Pratim Dutta +22Low-Resource MTMachine Translation

  6. MICRO: Multi-Fidelity Active Search for Severe Error Discovery

    Sep 22, 2026Orlando Leone, Niclas Pokel, Pehuén Moure +2Human-in-the-Loop AnnotationMT Evaluation

  7. LocQE: Principled Domain Adaptation for Localisation Quality Estimation by Leveraging Post-Edits

    Sep 16, 2026Kathy Hämmerl, Gabriel Bretschner, Joern WuebkerDomain AdaptationMachine Translation

  8. TACTICS: Taxonomy-Aware Intelligent Corpus Sampling for Machine Translation

    Sep 16, 2026Prasanth Bathala, Anubhav Shrimal, Sukhdeep Singh Kharbanda +2Adaptive SamplingMachine Translation

  9. TransClean: A Benchmark for Detecting and Extracting Clean Translations from Large Language Model Outputs

    Sep 11, 2026Shenbin Qian, Yves ScherrerMachine TranslationMT Evaluation

  10. Improving Term Evaluation in Machine Translation: Variation Matters

    Sep 8, 2026Nicolas Dahan, Ziqian Peng, François Yvon +1Machine TranslationMT Evaluation

  11. Last Translation Benchmark

    Sep 3, 2026Vilém Zouhar, Niyati Bafna, Mukund Choudhary +257MT Evaluation

  12. Beyond BLEU: A Case for Redefining Sign Language Translation Benchmarks

    Sep 3, 2026Oline Ranum, Edward Fish, Simon Hadfield +1Sign Language TranslationMT Evaluation

  13. IndicQE-APE: A Consolidated Benchmark for Quality Estimation and Automatic Post-Editing for Indic Languages

    Aug 17, 2026Diptesh Kanojia, Archchana Sindhujan, Sourabh Deoghare +14Low-Resource Language ProcessingMT Evaluation

  14. Cultivar: A Contrastive and Locale-Oriented Translation Benchmark for Investigating Contamination and Localisation Robustness

    Aug 10, 2026Pinzhen Chen, Koel Dutta Chowdhury, Xiaoya Xu +20Data Contamination in Language ModelsLanguage Model Robustness

  15. IDRAAK: From Multi-Agent NLP to Few-Shot Prompting for Semantic Drift Detection in Technical Requirements

    Aug 9, 2026Shiva AhirRequirements EngineeringFew-Shot Prompting

  16. Do Evaluation Metrics Detect Errors in Classical Chinese to English Translations?

    Aug 8, 2026Osvaldo Quinjica, Eric Bennett, Xinchen Yang +2Machine TranslationMT Evaluation

  17. Towards End-to-End Multilingual Metaphor Processing: Integrating Detection, Translation, and Evaluation

    Aug 4, 2026Jiahui Liang, Lifeng HanFigurative Language UnderstandingMachine Translation

  18. Looking under the Wrong Lamppost: On the Limitations of Automated Translation Quality Estimation

    Aug 4, 2026Serge Gladkoff, Angelika Vaasa, Sue Ellen Wright +2Automated EvaluationMT Evaluation

  19. Predicting Multilingual Classification and Translation Performance of LLMs with Cross-Lingual Alignment \unicodex2013\unicode{x2013} Is English Enough?

    Aug 4, 2026Adnan Al Ali, Kathy Hämmerl, Jindřich Libovický +1Multilingual Language Model EvaluationCross-Lingual Representation Alignment