---
title: Diffractive Optical Processors
url: https://www.emergentmind.com/topics/diffractive-optical-processors
type: topic
---

# Diffractive Optical Processors

Diffractive optical processors are optical computing and imaging systems in which one or more spatially engineered diffractive surfaces transform an incident electromagnetic field into a desired output field or intensity distribution through free-space propagation, interference, and detection. In the recent literature, they appear under closely related labels such as diffractive optical processors, diffractive optical networks, diffractive deep neural networks, and diffractive neural networks. Reported implementations span passive phase-only stacks, complex-amplitude metasurfaces, reconfigurable spatial-light-modulator platforms, and hybrid optical-digital systems jointly optimized for imaging, sensing, linear transforms, encryption, and machine learning tasks [2406.10688][2412.11374].

## 1. Physical principles and forward models

The standard forward model treats each diffractive layer as a thin transmissive mask and the spaces between layers as free-space propagation segments. In angular-spectrum form, if \(U_l(x,y;\lambda)\) denotes the optical field after layer \(l\), propagation over distance \(z\) is commonly written as
$$
U_{l+1}(x,y;\lambda)=\mathcal{F}^{-1}\left\{\mathcal{F}[U_l(x,y;\lambda)]\cdot H(f_x,f_y;\lambda,z)\right\},
$$
with
$$
H(f_x,f_y;\lambda,z)=\exp\!\left[i\,2\pi z\sqrt{1/\lambda^2-f_x^2-f_y^2}\right],
$$
or, in equivalent Rayleigh–Sommerfeld form, as a convolution with the free-space kernel [2412.11374]. Closely related formulations are used in monochrome linear-transform processors, multiplexed permutation devices, hybrid optical-digital systems, and complex-field imagers [2512.06658][2402.02397][2406.10688][2401.16779].

Layer modulation is most often phase-only, with transmission written as \(t_\ell(x,y)=\exp[j\phi_\ell(x,y)]\), although amplitude-phase parameterizations of the form \(t_\ell(x,y)=a_\ell(x,y)\exp[i\psi_\ell(x,y)]\) are also reported [2212.12873][2512.06658]. In fabricated devices, the trainable variable is usually local thickness, which sets the phase delay through the material refractive index and, when absorption is non-negligible, the amplitude response as well [2401.16779][2311.04473]. This makes the optical processor a cascaded linear operator in the complex field domain.

A persistent conceptual point is that linearity depends on what is treated as the signal. Many diffractive processors are linear with respect to the complex optical field, but the measured output is often an intensity, \(I_{\rm out}=|U_{\rm out}|^2\), which introduces a nonlinear mapping between input representation and detector readout [2212.12873]. Under spatially incoherent illumination, the processor is described instead by an intensity point-spread function \(H(m,n;m',n')=|h(m,n;m',n')|^2\), yielding
$$
I_{\rm out}(m,n)=\sum_{m',n'} |h(m,n;m',n')|^2\,I_{\rm in}(m',n'),
$$
so the diffractive volume acts as a linear operator on time-averaged intensity rather than on coherent field amplitudes [2303.13037].

## 2. Inverse design, optimization, and robustness

Recent diffractive optical processors are typically obtained by end-to-end inverse design. The optical forward model is embedded in an autodifferentiable framework, and layer parameters are optimized by backpropagation using task-dependent losses. Reported software stacks include PyTorch, TensorFlow, and direct in situ optimization on physical hardware [2412.11374][2203.13482][2507.05583].

Training objectives vary with the task. Imaging systems use MSE, NMSE, PCC, SSIM-related terms, or diffraction-efficiency penalties; classification systems use cross-entropy or detector-energy losses; linear-transform processors use output-field or transformation-matrix MSE; and hybrid systems jointly update optical and digital parameters with a common end-to-end loss [2412.11374][2512.06658][2409.08423][2506.03317]. A representative visible unidirectional imager combined NMSE, PCC, and energy-throughput terms over both forward and backward directions,
$$
L=\sum_\lambda \left[\alpha\cdot \mathrm{NMSE}(O,I)-\beta\cdot \mathrm{PCC}(O,I)+\gamma\cdot \eta_{\rm diff}\right],
$$
with random wavelength sampling from red, green, and blue bands to enforce broadband operation [2412.11374].

Robustness to fabrication and alignment errors is a recurring design requirement. Several works inject random axial and lateral shifts during training, a procedure often termed “vaccination,” to improve tolerance to misalignment, fabrication error, and lifetime drift [2412.11374][2212.12873][2311.04473]. A distinct calibration strategy appears in the reconfigurable diffractive processing unit, where in-silico pre-training is followed by layer-wise adaptive fine-tuning using experimentally measured intermediate fields to compensate aberration, misalignment, and non-ideal modulator response [2008.11659].

When accurate physical modeling is difficult, reported in situ learning methods optimize the hardware directly. A model-free reinforcement-learning approach based on Proximal Policy Optimization treats the optical system as an environment, updates phase patterns from measured rewards, reuses each in situ batch for multiple gradient steps, and experimentally shows better convergence and performance in tasks including energy focusing through a random diffuser, holographic image generation, aberration correction, and optical image classification [2507.05583]. This suggests that diffractive optical processors increasingly occupy a continuum between fully modeled inverse design and hardware-in-the-loop optimization.

## 3. Expressivity, universality, multiplexing, and reconfiguration

A major theoretical theme is the representation of arbitrary linear operators. Under spatially coherent light, a phase-only diffractive network can implement arbitrary complex-valued linear transformations between input and output fields-of-view when the total number of trainable diffractive features satisfies \(N\ge 2N_iN_o\); under spatially incoherent monochromatic light, the same scaling is reported for arbitrary linear intensity transformations [2303.13037]. For polarization-multiplexed diffractive computing, the reported feature-count scaling is \(N\gtrsim N_pN_iN_o\), where \(N_p\) is the number of distinct transformations assigned to input-output polarization pairs [2203.13482]. For illumination phase multiplexing, a monochrome diffractive network is reported to realize \(T\) distinct complex-valued linear transformations with \(N\ge 2TN_iN_o\), with numerical demonstration of \(T=512\) transformations at negligible error [2512.06658].

These results connect to a broader family of multiplexing schemes. Reported mechanisms include polarization, wavelength, bidirectional propagation, mechanical rotation of layers, and input phase diversity. A mechanically reconfigurable \(K\)-layer permutation processor performs up to \(4^K\) independent permutation operations because each layer can take four discrete orientations \(\{0^\circ,90^\circ,180^\circ,270^\circ\}\) [2402.02397]. A bilayer cascaded-metasurface classifier uses wavelength and polarization degrees of freedom to perform dual-task or tri-task recognition on MNIST, FMNIST, and KMNIST, with tri-task accuracies remaining greater than \(80\%\) and improving under end-to-end joint optimization of physically realizable meta-atoms [2409.08423]. An all-optical autoencoder exploits bidirectional multiplexing so that the same stack functions as an encoder in one propagation direction and as a decoder in the opposite direction, defining a diffractive latent space with compression ratios of approximately \(52\times\) or \(64\times\), depending on latent geometry [2409.20346].

| Multiplexing or reconfiguration mechanism | Reported capability | Paper |
|---|---|---|
| Polarization encoding | Multiple arbitrary linear transforms with \(N\gtrsim N_pN_iN_o\) | [2203.13482] |
| Illumination phase keys | \(T=512\) complex-valued transforms in one monochrome network | [2512.06658] |
| Layer rotations | Up to \(4^K\) permutation operations | [2402.02397] |
| Wavelength/polarization metasurfaces | Dual-task and tri-task classification | [2409.08423] |
| Bidirectional propagation | Optical encoding and decoding in one stack | [2409.20346] |

A common misconception is that a fabricated diffractive stack can perform only one fixed function. The reported literature shows instead that a single physical processor can execute multiple transformations when distinct channels or control variables are available, such as input polarization, wavelength, illumination phase profile, mechanical rotation state, or propagation direction [2203.13482][2512.06658][2402.02397][2409.20346].

## 4. Imaging, wavefront processing, and computational sensing

Diffractive optical processors have been used to implement a wide range of imaging and sensing modalities. A two-layer visible-spectrum unidirectional imager fabricated on high-purity fused silica forms high-fidelity images in the forward direction while generating weak, distorted patterns in the backward direction. Over 200 test wavelengths from 450 to 650 nm, the reported two-layer design achieved forward PCC \(>0.86\), backward PCC \(<0.58\), forward diffraction efficiency \(>28\%\), and backward diffraction efficiency \(<13\%\); a three-layer variant improved these figures to forward PCC \(>0.89\), backward PCC \(<0.33\), forward efficiency \(>30\%\), and backward efficiency \(<10\%\) [2412.11374].

Several reported systems target direct optical recovery of information that conventional intensity sensors do not natively provide. A complex-field imager uses successive diffractive surfaces to create two output channels that perform amplitude-to-amplitude and phase-to-intensity transformations, thereby enabling snapshot imaging of both amplitude and quantitative phase without digital reconstruction, within an axial span of approximately \(100\) wavelengths [2401.16779]. A multispectral quantitative phase imaging processor uses 10 phase-only layers to encode phase profiles at 9 or 16 visible wavelengths into spatially separated intensity patterns on a monochrome focal-plane array, yielding \(17.04\pm0.33\) dB PSNR and \(0.770\pm0.015\) SSIM for the 9-channel design, and \(16.67\pm0.43\) dB PSNR and \(0.726\pm0.031\) SSIM for the 16-channel design on unseen MNIST-based phase objects [2308.02952]. A solid-immersion diffractive optical processor couples a high-index encoder to decoder layers in air to resolve subwavelength phase and amplitude features; the reported terahertz proof of concept experimentally resolved \(w\approx0.293\lambda\approx0.22\) mm, described as sub-Rayleigh by approximately \(22\%\) [2401.08923].

Wavefront engineering is another major application. A diffractive optical phase-conjugation processor trained on random Zernike-aberrated inputs approximates the conjugate phase distribution and was experimentally validated in the terahertz regime. In the reported \(K=8\) transmissive design, blind two-mode tests yielded phase MAE \(=1.38\pm0.12\%\) and amplitude MAE \(=8.89\pm1.91\%\); depth also improved diffraction efficiency, from approximately \(13.8\%\) at \(K=4\) to approximately \(24.4\%\) at \(K=10\) [2311.04473].

Diffractive processors also appear as optical preconditioners for difficult propagation environments. An interleaved diffractive network for information transfer through random diffusers inserts trainable layers within a volumetric scattering medium and reports PCC improvement from approximately \(0.65\) at \(K=2\) to approximately \(0.8\) at \(K=5\) on unseen diffusers, with a best PCC of approximately \(0.9\) at the smallest interplane spacing of \(6.7\lambda\); a jointly trained hybrid system with a U-Net-style backend of approximately \(7.6\)k parameters further improves robustness to random rotations, shifts, and scaling [2603.07975]. In structural health monitoring, a single reflective diffractive layer jointly optimized with shallow neural networks remotely encodes 3D structural vibration spectra into four detector signals; the reported spectral MSE in the \([9,11]\) Hz band was \(1.11\times10^{-2}\) for the jointly optimized diffractive layer, compared with \(1.42\times10^{-1}\) for a separately optimized diffractive layer, \(3.58\times10^{-1}\) for a Fresnel lens array, and \(6.24\times10^{-1}\) for a random diffuser [2506.03317].

The same general framework has also been used for class-specific all-optical encryption and decryption, permutation-based encryption, and diffractive latent-space processing for denoising, classification, and image generation [2212.12873][2402.02397][2409.20346]. This suggests that “imaging” in the diffractive-processor literature often includes learned transforms that mix sensing, coding, inference, and optical encryption rather than conventional image formation alone.

## 5. Fabrication, materials, and hardware embodiments

Fabrication strategies span wafer-scale nanofabrication, metasurface manufacturing, two-photon polymerization, stereolithography, PolyJet printing, and programmable optoelectronic platforms. A visible-spectrum unidirectional imager was fabricated on a 6-inch high-purity fused silica wafer using 4× projection lithography and Cl\(_2\) dry etching to realize 16 discrete phase levels with approximately \(100\) nm step size. The process achieved lateral registration below \(3\,\mu\)m between front and back surfaces, \(3\)-\(5\%\) etch-depth error across the wafer, and throughput of approximately \(918\) diffractive-processor chips per wafer, totaling approximately \(0.5\) billion diffractive features [2412.11374].

In metasurface implementations, the trainable optical response is tied to a library of unit cells. A reported multi-task classifier uses TiO\(_2\) nanofins on glass with fixed height \(600\) nm, lattice period \(400\) nm, and trainable in-plane widths \(W_x,W_y\in[60\,\mathrm{nm},350\,\mathrm{nm}]\) to realize wavelength- and polarization-dependent complex transmittances derived from full-wave simulations at \(450\), \(550\), and \(650\) nm [2409.08423]. This differs from phase-only free-space stacks, but it remains within the diffractive-processor paradigm because cascaded diffraction and trainable local transmission still define the end-to-end operator.

Terahertz proof-of-concept systems frequently use 3D printing because the larger wavelength relaxes fabrication tolerances. Reported examples include class-specific encryption networks fabricated by two-photon polymerization in IP-Dip photoresist and tested at \(1550\) nm [2212.12873]; monolithic solid-immersion encoder-decoder pairs printed in VeroBlackPlus for terahertz subwavelength imaging [2401.08923]; stereolithography-fabricated phase-conjugation processors in isotropic polymer with \(n\approx1.7\) and \(k\approx0.017\) at \(\lambda=0.75\) mm [2311.04473]; and complex-field imagers, rotation-multiplexed permutation devices, and all-optical autoencoders validated with 3D-printed diffractive layers in the terahertz band [2401.16779][2402.02397][2409.20346].

A separate hardware lineage emphasizes programmable rather than fixed optics. The reconfigurable diffractive processing unit combines a digital micromirror device for amplitude encoding, a phase-only spatial light modulator with 8-bit phase resolution and \(1920\times1080\) pixels for synaptic weights, and an sCMOS detector array for optical readout and square-law nonlinearity. By time-multiplexing layers, the system supports diffractive feedforward and recurrent neural networks, experimentally achieving \(96.0\%\) MNIST accuracy for a 3-layer D2NN after adaptive training and video-accuracy up to \(100\%\) on Weizmann with a D-RNN++ configuration [2008.11659].

Integration with conventional imaging and photonic hardware is an explicit design goal in several reports. The visible unidirectional processor has an axial thickness of approximately \(2\) mm and lateral aperture of approximately \(366\,\mu\)m per processor, described as compatible with CMOS image sensors with \(1.12\,\mu\)m pixels, and potential on-chip integration includes silicon-photonic waveguide coupling and monolithic hybrid optoelectronic modules [2412.11374].

## 6. Linearity, nonlinear computation, limitations, and future directions

Diffractive optical processors are often described as linear optical systems, but the literature now distinguishes several routes to nonlinear computation. One route uses square-law detection together with suitable input encoding. Under phase encoding of scalar or vector inputs, a passive phase-only diffractive processor can approximate bandlimited nonlinear functions at the output intensity. A reported framework establishes universal approximation for arbitrary sets of bandlimited nonlinear functions, including multivariate and complex-valued functions, and numerically demonstrates one million distinct nonlinear functions computed in parallel by a diffractive processor [2507.08253]. A closely related incoherent-light framework uses intensity-only encoding and differential detector readout to realize universal nonlinear approximation under spatially incoherent or partially coherent illumination, again with numerical results showing snapshot computation of up to one million distinct nonlinear functions in a single forward pass [2603.29131].

A second route introduces explicit input-dependent optical modulation. The recurrent diffractive optical neural processor, ReDON, senses a fraction of the propagating field, processes it with a lightweight parametric function, and uses the result for electro-optic self-modulation in later layers. This reconfigurable self-modulated nonlinearity is combined with recurrence through hardware reuse of a fixed passive metasurface stack. On image recognition and segmentation benchmarks, the reported gains reach up to \(20\%\) in test accuracy or mIoU relative to prior diffractive optical neural networks with comparable model complexity, while the self-modulation overhead remains below \(1\) mW compared with more than \(100\) mW laser power [2602.23616].

A third route is hybrid optical-digital inference. The survey on programmable diffraction describes such systems as establishing a “diffractive language” between analog wave processing and digital neural networks, and reports applications spanning classification, computational imaging, single-pixel sensing, and programmable metasurface sensing [2406.10688]. In individual case studies, jointly optimized optical front ends and shallow or compact digital back ends improve reconstruction fidelity, classification accuracy, or task robustness while keeping the optical front end passive [2506.03317][2603.07975].

Several limitations recur across the literature. Phase-only modulation is widely preferred in practice but doubles the degrees-of-freedom requirement relative to full complex modulation for universal linear transforms [2303.13037]. Static fabricated processors can require redesign or refabrication to change target functions unless reconfiguration is introduced through SLMs, phase keys, multiplexing, or mechanical rotation [2203.13482][2512.06658][2402.02397]. Alignment sensitivity, fabrication error, hardware drift, and model mismatch remain central concerns, motivating vaccination during training, adaptive experimental fine-tuning, and model-free in situ learning [2412.11374][2008.11659][2507.05583].

A further misconception is that diffractive optical processors are only useful under fully coherent laser illumination. Reported systems operate under coherent, partially coherent, and spatially incoherent conditions, provided the forward model and training objective are matched to the source statistics and detection physics [2406.10688][2303.13037][2603.29131]. A plausible implication is that the field is moving from narrowly defined optical neural networks toward a broader class of learned wave processors in which fabrication method, illumination coherence, multiplexing strategy, and optical-digital partitioning are all design variables rather than fixed assumptions.

Source: https://www.emergentmind.com/topics/diffractive-optical-processors