How Language Directions Align with Token Geometry in Multilingual LLMs

Multilingual LLMs demonstrate strong performance across diverse languages, yet there has been limited systematic analysis of how language information is structured within their internal representation space and how it emerges across layers. We conduct a comprehensive probing study on six multilingual LLMs, covering all 268 transformer layers, using linear and nonlinear probes together with a new Token--Language Alignment analysis to quantify the layer-wise dynamics and geometric structure of language encoding. Our results show that language information becomes sharply separated in the first transformer block (+76.4$\pm$8.2 percentage points from Layer 0 to 1) and remains almost fully linearly separable throughout model depth. We further find that the alignment between language directions and vocabulary embeddings is strongly tied to the language composition of the training data. Notably, Chinese-inclusive models achieve a ZH Match@Peak of 16.43\%, whereas English-centric models achieve only 3.90\%, revealing a 4.21$\times$ structural imprinting effect. These findings indicate that multilingual LLMs distinguish languages not by surface script features but by latent representational structures shaped by the training corpus. Our analysis provides practical insights for data composition strategies and fairness in multilingual representation learning. All code and analysis scripts are publicly available at: https://github.com/thisiskorea/How-Language-Directions-Align-with-Token-Geometry-in-Multilingual-LLMs.

Paper

References (11)

07UnsupervisedCross-lingualRepresentationLearning at Scale2020 · Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics
09How Multilingual is Multi-lingual BERT?2019 · Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics
102024. Qwen2.5 Technical ReportarXiv preprint
112024.TheLlama3HerdofModelsarXiv preprint

Similar papers

© 2026 NYSGPT2525 LLC