---
title: Disorder-Enabled Joint Projection
url: https://www.emergentmind.com/topics/disorder-enabled-joint-projection
type: topic
---

# Disorder-Enabled Joint Projection

Searching arXiv for the cited works and closely related terminology to ground the article in current papers.
arXiv search: "Disorder-enabled Synthetic Metasurfaces"
Disorder-enabled joint projection denotes a class of mappings in which disorder, heterogeneous allocation, or cross-view alignment is used to project multiple functions, signals, or representations into a shared physical or latent substrate. In the recent literature considered here, the term appears in three distinct settings: disordered photonic lattices that realize random projections satisfying Johnson–Lindenstrauss-type guarantees [2108.08654], synthetic metasurfaces that use engineered structural disorder to encode multiple optical functions in a single aperture and jointly project them in real and momentum space [2507.04696], and cross-view contrastive learning that aligns volumetric imaging and ROI-graph embeddings for brain disorder classification [2603.10253]. The common technical motif is not a single formalism but a shared objective: preserve task-relevant structure while jointly embedding multiple degrees of freedom into a compact projector.

## 1. Scope and terminological usage

In the cited literature, “disorder-enabled joint projection” is associated with three different projection targets: lower-dimensional vector embeddings, multifunctional optical fields, and shared latent representations for classification. This suggests a family resemblance rather than a single standardized doctrine.

| Domain | Projected entities | Projection mechanism |
|---|---|---|
| Disordered photonic lattices | Vectors in $\mathbb{C}^N$ | Row-sampled unitary evolution with intermediate diagonal disorder |
| Synthetic metasurfaces | Multiple optical functions | Disordered meta-pixel shares with qBIC-based selectivity |
| Brain disorder classification | Imaging and ROI embeddings | Bidirectional cross-view contrastive alignment |

A notable terminological distinction is that, in the photonic works, “disorder” refers to engineered physical randomness or structural nonuniformity, whereas in the neuroimaging work the term is tied to brain disorder classification rather than to physical disorder [2603.10253]. The shared word “projection” is likewise domain-specific: in photonic lattices it denotes a dimensionality-reducing linear map, in metasurfaces a far-field reconstruction from selected meta-pixel subsets, and in neuroimaging a learned map into a common latent space.

## 2. Random complex projections in disordered photonic lattices

In the photonic-lattice formulation, one considers an array of $N$ single-mode waveguides with nearest-neighbor evanescent coupling and diagonal disorder. The coupled-mode dynamics are

$$
\frac{d}{dz}\mathbf{u}(z) \;=\; -\,i\,H\,\mathbf{u}(z),
\qquad
\mathbf{u}=(u_1,\dots,u_N)^T,
$$

with tight-binding Hamiltonian

$$
H_{ij} \;=\; \begin{cases}
\delta_i, & i=j,\\
\kappa,    & |i-j|=1\text{ (nearest neighbors)},\\
0,         & \text{otherwise.}
\end{cases}
$$

Here $\kappa$ is the uniform coupling constant, and each detuning $\delta_i$ is drawn i.i.d. from a zero-mean distribution of width $\sigma$. Over propagation length $L$, the evolution operator is the unitary matrix

$$
U \;=\;\exp\!\bigl(-i\,H\,L\bigr)\;\in U(N).
$$

In the eigenmode basis, the matrix elements satisfy

$$
U_{ij} \;=\; \sum_{n=1}^N e^{-i\,\beta_n L}\;\phi^{(n)}_i\,\phi^{(n)*}_j.
$$

When the disorder is neither too weak nor too strong, the sum over modes becomes well mixed, and the entries of $U$ approach circular complex Gaussians:

$$
\Re U_{ij},\;\Im U_{ij} \;\overset{\text{d}}{\longrightarrow}\; \mathcal{N}\bigl(0,\tfrac1{2N}\bigr)
\quad \Longrightarrow \quad
U_{ij}\sim\mathcal{CN}\!\bigl(0,1/N\bigr).
$$

This Gaussianization is the key step that connects physical propagation to random projection theory [2108.08654].

The resulting dimensionality reduction is implemented by row-sampling the unitary. If $\bar U$ is the $M\times N$ matrix formed by selecting $M$ output channels out of $N$, then $\mathbf{x}\mapsto \bar U\,\mathbf{x}$ acts as a partial-unitary embedding. The cited formulation gives a complex Johnson–Lindenstrauss statement: for a set of $n$ vectors in $\mathbb{C}^N$, if $R$ has i.i.d. $\mathcal{CN}(0,1/M)$ entries and

$$
M \;\ge\; \frac{8\ln n + 2\ln(1/\delta)}{\varepsilon^2},
$$

then with probability at least $1-\delta$ all pairwise distances are preserved within a factor $(1\pm\varepsilon)$. Under the approximation $U_{ij}\approx\mathcal{CN}(0,1/N)$ and near-orthonormal sampled rows, the same argument applies to $\bar U$. For any fixed $\mathbf{v}\in\mathbb{C}^N$,

$$
\Pr\Bigl[\bigl|\|\bar U\,\mathbf{v}\|_2^2 - \|\mathbf{v}\|_2^2\bigr|\ge \varepsilon\,\|\mathbf{v}\|_2^2\Bigr]
\;\le\;
2\exp\!\Bigl(-\tfrac{\varepsilon^2}{4}\,M\Bigr),
$$

so the embedding dimension scales as $M=O(\varepsilon^{-2}\ln(n/\delta))$ [2108.08654].

## 3. Disorder regime, transport physics, and implementation constraints

The same work distinguishes three transport regimes as $\sigma/\kappa$ increases. In the ballistic regime, $\sigma\ll\kappa$, disorder is too weak: eigenstates are extended plane waves, phases remain highly structured, and the matrix entries are neither zero-mean nor Gaussian. In the localized regime, $\sigma\gg\kappa$, Anderson localization dominates, each input excites only a few nearby sites, and the columns of $U$ become sparse. The desired projection behavior therefore requires an intermediate diffusive regime, $\sigma\sim\kappa$, in which wave interference over many scattering events yields random-matrix-type mixing [2108.08654].

The stated conditions for Johnson–Lindenstrauss behavior are

$$
1/\sigma \;\ll\; L
\quad\text{and}\quad
\xi(\sigma)\;=\;\mathcal{O}\!\bigl((\kappa/\sigma)^2\bigr)\;\gg\;L,
$$

together with the practical inequality

$$
5/\sigma \;\le\; L\;\le\; 0.2\,\xi(\sigma).
$$

The article further reports that one may sweep $\sigma\in[0.2\kappa,2\kappa]$ and verify Gaussianity of $\{U_{ij}\}$ באמצעות a Kolmogorov–Smirnov test. Since $\mathbb{E}\|\bar U\mathbf{x}\|^2=\tfrac{M}{N}\|\mathbf{x}\|^2$, the output is rescaled by $\sqrt{N/M}$ to compensate average power loss [2108.08654].

Several operational modes are described. In time-multiplexing, a train of $n$ short pulses carries the inputs $\mathbf{x}_i$, and the $i$th time bin at the output contains $\bar U\,\mathbf{x}_i$. In multi-channel excitation, disjoint input waveguides or orthogonal modulation codes suppress cross-talk when signals are sparse in time or wavelength. Correlated inputs remain compatible with the guarantee because the Johnson–Lindenstrauss bound is uniform over all pairs; clustered inputs remain clustered after projection. A common misconception, explicitly contradicted by the regime analysis, is that increasing disorder monotonically improves randomness. The cited result instead requires intermediate diagonal disorder: too little disorder fails to randomize, while too much induces localization [2108.08654].

## 4. Engineered disorder in synthetic metasurfaces

The metasurface formulation treats disorder as a design variable that enables many optical functions to coexist within a single aperture. The proposed optimization balances per-function fidelity with disorder regularization:

$$
J\;=\;\sum_{m=1}^M
\bigl\|\,S_m(\mathbf{r})-T_m(\mathbf{r})\bigr\|^2
\;+\;\lambda\,R\bigl(\{\mathbf{r}_i\}\bigr),
$$

where $S_m(\mathbf{r})$ is the simulated field or phase distribution for function $m$, $T_m(\mathbf{r})$ is the target profile, $\{\mathbf{r}_i\}$ are the meta-pixel coordinates, and $R(\{\mathbf{r}_i\})$ penalizes too much clustering or too much ordering. The regularizer is given as

$$
R=\sum_{i\neq j}\exp\!\Bigl[-\frac{|\mathbf{r}_i-\mathbf{r}_j|^2}{2\sigma^2}\Bigr]
\;-\;\alpha\,\sum_{i\neq j}\exp\!\Bigl[-\frac{|\mathbf{r}_i-\mathbf{r}_j|^2}{2\Sigma^2}\Bigr].
$$

Varying $\alpha$, $\sigma$, and $\Sigma$ tunes the pattern between minimum-clustering and overly sparse arrangements, while $\lambda$ balances function fidelity against disorder quality [2507.04696].

Spectral selectivity is implemented through nonlocal meta-pixels engineered to support quasi-bound states in the continuum. In an idealized lossless structure, a BIC has $Q\to\infty$; slight symmetry breaking yields a qBIC resonance with

$$
Q\;\sim\;10^2\!-\!10^4
\quad\Longrightarrow\quad
\Delta\lambda\;=\;\frac{\lambda}{Q}\;\le\;1\text{–}2\,\mathrm{nm}.
$$

The transmission coefficient is modeled as

$$
t_i(\lambda)\;\approx\;
\exp\bigl[i\,\phi_i(\lambda)\bigr]\,
\frac{\Gamma/2}{\,\omega_i-\omega-i\,\Gamma/2\,},
$$

with $\Gamma=\omega_i/Q$. By rotating each T-shaped meta-pixel in-plane, the geometric phase $\phi_i\in[0,2\pi)$ can be set independently without degrading the $Q$-factor [2507.04696].

A central empirical claim is that a single optical function does not require a contiguous aperture. Random sub-sampling to a fraction $p$ yields a functional density $D_f=1/p$. Using the angular-spectrum method numerically and the Strehl ratio experimentally, the study reports that disordered sampling maintains $SR>0.9$ down to $p\approx0.1$, whereas an ordered sector-shaped aperture with the same $p$ already fails, reaching $SR<0.5$ by $p\approx0.5$. Disorder is quantified using Moran’s index,

$$
I_m=\frac{\sum_{i\neq j} w_{ij}(r_i-\bar r)(r_j-\bar r)}
{\sum_{i\neq j} w_{ij}\,(r_i-\bar r)^2},
$$

and the Strehl ratio is reported to recover suddenly as $I_m$ drops from $1$ to $0$. The filling factor is also constrained by a spatial Nyquist condition,

$$
p\,\bigl(\Delta x \bigr)^2 \;\gtrsim\;\bigl(\lambda/2\rho_{\max}\bigr)^2,
$$

where $\rho_{\max}$ is the maximum spatial frequency of the wavefront. In the demonstrated platform, each meta-pixel is $8.1\,\mu\mathrm{m}\times 8.1\,\mu\mathrm{m}$, and at $p=1/11$ the design achieves $D_f=11$ while avoiding Bragg diffraction orders [2507.04696].

## 5. Real-space and momentum-space joint projection

The metasurface proof of concept combines spectral multiplexing of lens profiles with polarization multiplexing of gratings. For the spectral channel, $M=11$ wavelengths $\{\lambda_i\}\subset[1200,1400]\,\mathrm{nm}$ are each assigned a randomly placed subset of meta-pixels. A pixel at position $\mathbf{r}_i$ carries a resonance tuned to $\lambda_i$ and a rotation encoding the lens phase

$$
\phi_i(\mathbf{r}_i;\lambda_i)
\;=\;-\,k_i\!\Bigl(\sqrt{|\mathbf{r}_i|^2+f^2}-f\Bigr),
\qquad
k_i=2\pi/\lambda_i.
$$

Because of the high $Q$-factor, the transmission is approximated by

$$
t_i(\lambda)=
\begin{cases}
e^{i\phi_i} & \text{if }\lambda\approx\lambda_i,\\
0 & \text{otherwise.}
\end{cases}
$$

Subsampling to $p=1/11$ for each wavelength leaves the remaining area available for the other ten functions [2507.04696].

For polarization multiplexing, three disordered shares, each with $p=1/3$, are assigned to momentum-space gratings corresponding to the orthogonal bases horizontal/vertical, diagonal/anti-diagonal, and right/left circular. On resonance for polarization basis $\pi$, the transmission is

$$
t_i(\lambda_0,p=\pi)\;=\;e^{i\,\mathbf{g}_\pi\cdot\mathbf{r}_i},
$$

with $\mathbf{g}_\pi$ the grating wave vector; for example, H is deflected to $+\mathbf{g}_{HV}$ and V to $-\mathbf{g}_{HV}$. The full far-field intensity is

$$
I(\theta,\phi;\lambda,p)
\;=\;
\biggl|\sum_{i}t_i(\lambda,p)\,
e^{\,i\,\mathbf{k}(\theta,\phi)\cdot\mathbf{r}_i}
\biggr|^2.
$$

When wavelength and polarization select a single share, the other contributions vanish approximately and the target lens or grating is reconstructed in one shot [2507.04696].

The reported demonstration includes a synthetic achromatic metalens with aperture diameter $8.1\,\mathrm{mm}$, focal length $f\approx10\,\mathrm{mm}$, and $\mathrm{NA}\approx0.4$, corresponding to a diffraction-limited spot of approximately $\lambda/(2\mathrm{NA})\approx1.5\,\mu\mathrm{m}$. The measured Strehl ratio is $\overline{SR}=0.87\pm0.07$ over $1200$–$1400\,\mathrm{nm}$, and the chromatic focal shift is $\lesssim1/40$ of $f$, compared with a $2\,\mathrm{mm}$ shift for a single-wavelength reference lens. The experimentally measured resonance quality factor is $Q\approx150$, while simulations and deeper patterning can reach $Q\sim10^3$ [2507.04696].

The same platform supports single-shot polarimetric imaging with minimum super-pixel size $3.2\,\mu\mathrm{m}\times 3.2\,\mu\mathrm{m}$, spatial resolution of approximately $3\,\mu\mathrm{m}$, and polarization reconstruction error $0.039\pm0.017$. The work reports imaging of radially and azimuthally polarized vector beams, whose raw $k$-space spots appear as six lobes mapping exactly to H/V, D/A, and R/L content, as well as characterization of an optical skyrmion with topological number approximately $0.99$ in a single acquisition [2507.04696].

The same framework is explicitly extended in the cited discussion to arbitrary hologram multiplexing, OAM holography using helical phases $e^{i\ell_m\varphi}$, and higher-dimensional joint control over wavelength, polarization, OAM, angle of incidence, and temporal waveforms. Trade-offs are also stated: cross-talk increases when resonances overlap, higher $Q$ tightens spectral isolation but enlarges pixel footprints and slows response, disorder-induced speckle or side-lobes can be reduced with a band-limited regularizer, and fabrication tolerances may be compensated by in-line metrology and neural-network-based correction [2507.04696].

## 6. Cross-view joint projection for brain disorder classification

In neuroimaging, Liang and He describe a joint imaging-ROI representation learner in which volumetric and graph-based subject representations are projected into a shared latent space through bidirectional contrastive alignment [2603.10253]. The imaging branch takes a preprocessed $3$D T1-weighted MRI volume $x_i\in\mathbb{R}^{H\times W\times D}$ and applies a 3DSC-TF encoder, described as a hybrid 3D depthwise-separable CNN plus Transformer, to produce a $d_{\mathrm{img}}$-dimensional embedding. The ROI branch takes a subject-specific graph $G_i=(V,E)$, where $V$ are AAL parcel mean intensities and $E$ are Pearson correlations, and uses a GNN called “NeuroGraph” with $L=3$ message-passing layers and hidden size $d_{\mathrm{hidden}}$ to produce a $d_{\mathrm{roi}}$-dimensional embedding. Two projection heads, both two-layer MLPs, map these embeddings to a shared dimension $d_p$:

$$
z_{\mathrm{img}}^{(i)} = g_{\mathrm{img}}(x_i),\qquad
p_{\mathrm{img}}^{(i)} = h_{\mathrm{img}}(z_{\mathrm{img}}^{(i)}),
$$

$$
z_{\mathrm{roi}}^{(i)} = g_{\mathrm{roi}}(G_i),\qquad
p_{\mathrm{roi}}^{(i)} = h_{\mathrm{roi}}(z_{\mathrm{roi}}^{(i)}).
$$

For a mini-batch of size $B$, the similarity matrix is

$$
S_{ij} = \mathrm{sim}(p_{\mathrm{img}}^{(i)},\,p_{\mathrm{roi}}^{(j)})/\tau,
$$

with cosine similarity $\mathrm{sim}(u,v)=u^Tv/\|u\|\|v\|$ and temperature $\tau>0$. The bidirectional InfoNCE-style objective is

$$
L_{\mathrm{contra}}
= - \frac{1}{2B} \sum_{i=1}^B \left[
\log \frac{\exp(S_{ii})}{\sum_{j=1}^B \exp(S_{ij})}
+
\log \frac{\exp(S_{ii})}{\sum_{j=1}^B \exp(S_{ji})}
\right].
$$

Positive pairs are same-subject imaging and ROI embeddings, and all cross-subject pairs are negatives. After alignment, the projection heads are discarded and the base embeddings are concatenated,

$$
z_{\mathrm{fuse}}^{(i)} = [\,z_{\mathrm{img}}^{(i)}; z_{\mathrm{roi}}^{(i)}\,],
$$

then classified by an MLP plus softmax with cross-entropy loss. Training details reported in the cited description are AdamW with initial learning rate $1\mathrm{e}{-4}$, weight decay $1\mathrm{e}{-5}$, batch size $16$, up to $120$ epochs with early stopping on validation AUC, temperature $\tau=0.07$, contrastive weight $\lambda=1.0$ so that $L_{\mathrm{total}}=L_{\mathrm{cls}}+\lambda L_{\mathrm{contra}}$, dropout $0.3$ in the MLP heads, LayerNorm after each encoder block, end-to-end fine-tuning without separate pretraining, and $5$-fold stratified cross-validation with fixed folds [2603.10253].

Quantitatively, the joint model improves on both single-view baselines on ADHD-200 and ABIDE. On ADHD-200, ROI-only gives $\mathrm{Acc}=63.48\pm2.74\%$ and $\mathrm{AUC}=65.51\pm3.23\%$; imaging-only gives $\mathrm{Acc}=68.65\pm4.99\%$ and $\mathrm{AUC}=70.92\pm4.93\%$; the joint model gives $\mathrm{Acc}=69.29\pm4.44\%$, $\mathrm{AUC}=72.73\pm4.17\%$, and $\mathrm{F1}=69.01\pm4.51\%$. On ABIDE, ROI-only gives $\mathrm{Acc}=61.09\pm1.23\%$ and $\mathrm{AUC}=60.90\pm3.41\%$; imaging-only gives $\mathrm{Acc}=59.17\pm2.39\%$ and $\mathrm{AUC}=57.91\pm2.16\%$; the joint model gives $\mathrm{Acc}=62.54\pm1.79\%$, $\mathrm{AUC}=64.08\pm2.14\%$, and $\mathrm{F1}=61.71\pm1.43\%$. The summary statement in the source is that joint projection outperforms each single-view baseline by $1$–$2\%$ in Acc/AUC [2603.10253].

Interpretability analyses combine Grad-CAM on the imaging branch with saliency-based attribution on the ROI branch. Imaging-only maps are described as relatively diffuse and ROI-only maps as sharp but sometimes scattered across network edges. The joint model yields spatially coherent foci in the superior frontal gyrus, precentral gyrus, orbitofrontal cortex, and hippocampal or limbic regions. This suggests that the cross-view projection is functioning less as simple feature stacking than as a geometric alignment procedure that makes complementary global and local patterns mutually usable for the downstream classifier [2603.10253].

Source: https://www.emergentmind.com/topics/disorder-enabled-joint-projection