Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
A Paradigm Shift in Machine Translation: Boosting Translation Performance of Large Language Models
2024-01-16
· ICLR 2024 poster ·
anchor
Findings
IC-1236
Among 7B LLMs, Llama-2-7b achieves the best zero-shot COMET scores in both translation directions, while MPT-7b leads in BLEU for en-to-xx
IC-1237
Llama-2-7b's pre-existing translation knowledge is diluted by large amounts of parallel data, causing COMET to decline after 100k examples
IC-1238
Llama-2-13b produces off-target non-translation outputs in zero-shot English-to-foreign-language translation
IC-1239
Llama-2-7b and Llama-2-13b achieve top zero-shot cross-lingual performance among 7B models on XNLI, XStoryCloze, and XWinograd