Search papers, labs, and topics across Lattice.
This paper challenges the notion of a theoretical curse of multilinguality in embedding spaces by proving that the minimum dimensionality required for perfect multilinguality increases only logarithmically with the number of languages. The findings suggest that the observed degradation in performance as language coverage increases is not an inherent limitation of multilingual models, but rather a consequence of real-world data and training conditions. A small-scale empirical study supports this theoretical framework, providing new insights into multilingual NLP performance.
The curse of multilinguality may not be a fundamental barrier; instead, it鈥檚 a byproduct of data and training practices.
A central goal of multilingual NLP is to achieve high monolingual performance per language and cross-lingual alignment for large-scale language coverage with a multilingual model. The curse of multilinguality describes the phenomenon of degradation in multilingual model performance as we increase language coverage, posing a threat to the above goal. This paper asks whether multilingual embedding spaces are inherently incapable of achieving perfect multilinguality without a prohibitive increase in required capacity. We first formalize the goal of "perfect multilinguality", embodied in two multilinguality conditions. We then prove that the minimum dimensionality required for perfect multilinguality grows only logarithmically in the number of languages. That is, we show that there is no theoretical curse of multilinguality for embedding space structure. This suggests that the empirical curse of multilinguality is a result of real world data and training conditions. We back this understanding with a small-scale empirical study. Our paper provides the first theoretical and intrinsic perspective on the curse of multilinguality, with implications for the scientific understanding of this phenomenon.