Contrastive evaluation of pronoun translation in neural machine translation
-
Updated
Aug 22, 2019 - Perl
Contrastive evaluation of pronoun translation in neural machine translation
Python implementation of METEOR
This repository is for our paper "What do large language model need for machine translation evaluation?"
LLM translation evaluation dataset and report:Comparing DeepL, ChatGPT, Gemini & Claude on Chinese-to-English translation quality
Measuring Semantic Drift in Multilingual LLM-Generated Clinical Trial Explanations: A Reliability Framework
Grammatical Interpretable Scoring - An interpretable matric for Machine Translation Evaluation
Evaluate translations by either a self-hosted Embedder or using Chat-GPT as LLM-as-judge.
A reproducible evaluation ground for translation recipes, born from Remis.
To associate your repository with the translation-evaluation topic, visit your repo's landing page and select "manage topics."