Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
The Reasonableness Behind Unreasonable Translation Capability of Large Language Model
2024-01-16
· ICLR 2024 poster ·
anchor
Findings
IC-662
Translation ability in BLOOM models surges at approximately one-sixth of pre-training and then plateaus, with consistent dynamics across model sizes from 560M to 7.1B