Cross-Lingual Alignment Without Joint Training: Do Monolingual Language Models Converge on Universal Representations?
cs.CL
Submitted: 2026-08-27
Updated: 2026-08-27
Code: https://github.com/Ej-zhou/monolingual-alignment
License: http://creativecommons.org/licenses/by/4.0/
The gist: Cross-lingual alignment in multilingual language models is typically attributed to joint training: shared parameters, mixed-language batches, or explicit alignment objectives.
Terminology
Abstract
Cross-lingual alignment in multilingual language models is typically attributed to joint training: shared parameters, mixed-language batches, or explicit alignment objectives. We ask whether monolingual models trained on non-parallel data learn alignable representations without joint training. By testing on strictly monolingual language models, such as the Goldfish model families and independently developed models from different research labs, we find three results. Correlation: these models develop alignable representational geometry across layers, with alignment strengthening as data scale, model scale, or linguistic proximity increases. Construction: a single Procrustes rotation fit on parallel sentences maps hidden states between models. Causation: the same rotation transfers functional content; patching a rotated English residual into a German model on a factual cloze flips the prediction to the donor's capital in most cases. We confirm that cross-lingual alignment can emerge from the structure of language and the information it carries rather than from joint training, and this points to practical future directions including model stitching, merging, and modular multilingual systems built from monolingual components.
Sources
- Goldfish: Monolingual Language Models for 350 Languages
- Revisiting the Platonic Representation Hypothesis: An Aristotelian View
- Tracing Multilingual Representations in LLMs with Cross-Layer Transcoders
- BLOOM: A 176B-Parameter Open-Access Multilingual Language Model
- Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale
- A Systematic Assessment of Language Models with Linguistic Minimal Pairs in Chinese
- Exploiting Similarities among Languages for Machine Translation
- SGPT: GPT Sentence Embeddings for Semantic Search
- Bielik v3 Small: Technical Report
- Qwen2.5 Technical Report
- Gemma 2: Improving Open Language Models at a Practical Size
- Beyond the Final Layer: Intermediate Representations for Better Multilingual Calibration in Large Language Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering