Linear Separability of Language Representations
Determine whether each language occupies a distinct, linearly separable region of a multilingual Transformer’s activation space at every layer, such that a string generates text in language \(\mathcal{L}_i\) if and only if its representation at that layer lies in the corresponding region.
References
In the Linear Separability Hypothesis, each language \mathcal{L}_i is hypothesized to occupy a distinct region \mathcal{R}_ti of activation space for each layer t. Formally, a string s\in\Sigma* generates text in language \mathcal{L}_i if and only if its corresponding representations x_t\in\mathcal{R}_ti, where \mathcal{R}_ti is a linearly-separable region in embedding space.
The Semantic Hub Hypothesis provides crucial insight into how multilingual control should be designed. It posits that semantically equivalent inputs s{\mathcal{L}_s} and s{\mathcal{L}_t} from source and target languages have representations such that: