---
title: Quantum Block-Encoding Input Model
url: https://www.emergentmind.com/topics/block-encoding-input-model
type: topic
---

# Quantum Block-Encoding Input Model

The block-encoding input model is the mechanism by which a quantum algorithm receives matrix or operator data as a unitary-access primitive. In its standard form, a unitary \(U\) acting on ancillas and system is an \((\alpha,a,\varepsilon)\)-block-encoding of an operator \(A\) when \(\|A-\alpha(\langle 0^a|\otimes I)U(|0^a\rangle\otimes I)\|\le \varepsilon\); in the exact case, the top-left block of \(U\) is \(A/\alpha\). In recent work, the phrase “input model” denotes not only this definition but also the concrete assumptions under which such a unitary is compiled: sparse-access and state-preparation oracles, explicit arithmetic circuits, Fourier-symbol oracles, tensor-network contractions, and variationally trained unitaries all instantiate different block-encoding input models with sharply different normalization factors, ancilla counts, gate depths, and classical preprocessing requirements [2404.04554], [2301.08908], [2604.09538].

## 1. Formal role in quantum algorithms

Block encoding is the standard access primitive for QSVT, qubitization, and related polynomial-transform methods. In this framework, \(\alpha\) is the subnormalization factor, and the success amplitude in a postselection picture is \(1/\alpha\); when \(U\) acts on \(|0^a\rangle\otimes|\psi\rangle\), the ancilla-zero component is \((A/\alpha)|\psi\rangle\). In QSP- and QSVT-based simulation, the number of uses of the block-encoding scales linearly in \(\alpha t\) up to polylogarithmic factors in the target precision, so reducing \(\alpha\) directly reduces query complexity [2510.08644].

The same formalism accommodates non-Hermitian inputs. One standard route is Hermitian dilation, \(H(A)=\begin{bmatrix}0&A\\A^\dagger&0\end{bmatrix}\), which allows QSVT to act on singular values while preserving a Hermitian signal operator. This is used explicitly in quantum linear-algebra settings such as Kalman filtering, where addition, multiplication, and inversion are all expressed through block-encodings and composed inside a unified framework [2404.04554].

A central distinction in the literature is therefore between the abstract definition of a block-encoding and the concrete way the unitary is realized. Some constructions assume oracle access to entries or sparse structure; others compile the unitary explicitly from problem structure, so that no qRAM, signed amplitude loading, or black-box entry oracle is required. The latter trend is especially pronounced in recent work on differential operators, tensor networks, and many-body Hamiltonians [2604.09538].

## 2. Oracle-based, state-preparation, and arithmetic access models

A large class of input models is oracle-based. In dense-matrix settings, a common assumption is access to state-preparation oracles \(U_L\) and \(U_R\) that prepare row- and column-weighted superpositions; their composition yields a \((\|A\|_F,s,\varepsilon)\)-block-encoding of \(A\) with polylogarithmic query time in the matrix dimensions, provided the data are stored in a quantum-accessible structure [2404.04554]. In sparse and second-quantized settings, the input model is instead phrased in terms of a sparsity oracle \(O_C\) and an amplitude oracle \(O_A\). For second-quantized Hamiltonians, SWAP-based implementations of \(O_C\) and SELECT-SWAP data lookup for \(O_A\) reduce the T-count per oracle invocation to \(\tilde O(\sqrt{L})\) in the number of interaction terms \(L\) [2510.08644].

A different line of work replaces generic oracles by arithmetic descriptions of structure. For matrices with repeated values and patterned sparsity, one can specify nonzero entries by a value label \(d\), a multiplicity label \(m\), and reversible arithmetic maps between \((d,m)\) and row/column coordinates. In that model, the dominant data-loading cost depends on the number \(D\) of distinct values rather than the matrix dimension, and different schemes produce different subnormalizations: a base scheme with \(\alpha=\sqrt{S_cS_r}\,\|A\|_{\max}\), a preamplified scheme with \(\alpha=\sqrt{2}\,\mu_p(A)\), and a PREP/UNPREP scheme with \(\alpha_{1/2}=(\sqrt{S_cS_r}/D)\sum_d |A_d|\) when the requisite commutation conditions hold [2302.10949].

The dictionary-based sparse model pushes this idea further. There, a sparse matrix is organized into classes of repeated values \(A_l\), together with injective maps \(c_l(j)\) that specify where those values occur. The resulting unitary
\[
U_A=(\mathrm{UNPREP}\otimes I)\,O_c\,(\mathrm{PREP}\otimes I)
\]
block-encodes \(A\) with subnormalization \(\alpha=\sum_{l=0}^{s_0-1}|A_l|\), depth \(O(\log(ns))\), and ancilla count \(O(n^2s)\), where \(s\) is the number of nonzeros and \(s_0\le s\) is the number of dictionary items [2405.18007].

Approximate state-preparation approaches remain relevant for unstructured sparse data. S-FABLE block-encodes \(H^{\otimes n}AH^{\otimes n}\), then conjugates by outer Hadamards to recover a block-encoding of \(A\); LS-FABLE avoids the quadratic classical overhead by directly inserting scaled sparse entries into the rotation angles. For unstructured sparse matrices with \(O(N)\) nonzeros, the reported empirical behavior is approximately \(O(N)\) rotation gates and \(O(N\log N)\) CNOT gates after compression [2401.04234].

| Model | Access assumption | Characteristic feature |
|---|---|---|
| State-preparation oracles [2404.04554] | \(U_L,U_R\) for row/column amplitudes | Frobenius-norm normalization |
| Sparse oracle + amplitude oracle [2510.08644] | \(O_C,O_A\) | \(\tilde O(\sqrt{L})\) T-count |
| Arithmetic structured matrices [2302.10949] | reversible maps \(O_c,O_r,O_{\mathrm{data}}\) | loading depends on distinct values \(D\) |
| Dictionary sparse [2405.18007] | value classes and injective maps \(c_l(j)\) | \(\alpha=\sum_l |A_l|\), depth \(O(\log(ns))\) |
| S-FABLE / LS-FABLE [2401.04234] | classical angle preprocessing or sparse-entry access | aggressive circuit compression for sparse \(A\) |

## 3. Explicit structure-exploiting encodings for differential and Fourier operators

A prominent current direction is to eliminate generic data oracles entirely by exploiting operator structure. For the Difference-of-Gaussian operator on a periodic grid, the coefficients split into two normalized discrete Gaussian distributions \(p\) and \(q\). Exact Gaussian state-preparation circuits \(G_p\) and \(G_q\), a one-qubit branch indicator, a single Pauli-\(Z\) gate to encode the minus sign, and controlled cyclic shifts together yield an exact block-encoding of
\[
A_h=\sum_{t\in T}(p_t-q_t)S_t
\]
with \((\lambda,a,\varepsilon)=(2,s+1,0)\), independent of grid size \(N\), spatial dimension \(D\), and stencil width \(|T|\). The same work derives an exact success probability
\[
P_{\mathrm{succ}}=\frac14\sum_{\omega\in\mathbb Z_N^D}|\mu(\omega)|^2|\hat v_h(\omega)|^2
\]
and shows \(P_{\mathrm{succ}}=\Theta(h^4)\) for smooth inputs as the periodic grid is refined [2604.09538].

For finite-difference discretizations of the Laplacian on periodic grids, the input model is fully explicit: ancillas prepare fixed superpositions, and controlled cyclic shifts realize the stencil. In one dimension this gives \(\alpha=1\); in \(D\) dimensions the subnormalization is \(\alpha=D/2^{\lceil\log_2 D\rceil}\), with \(O(D\log N)\) T-gate complexity and \(p\sim h^4\) under \(C^4\) regularity assumptions [2509.02429].

QFT-based models access operators through their Fourier symbols. For bounded-domain fractional Laplacians with open, zero-extension boundary conditions, the native QFT implements a periodic circulant surrogate rather than the Toeplitz truncation. Zero-padding into an \(M\)-point periodic register and compressing back to the physical \(N\)-point subspace produces
\[
P_{N\to M}^\dagger \widetilde A_{\alpha,h}^{(M)} P_{N\to M}
=
A_{\alpha,h}^{(N)} + E^{(M)},
\]
where the residual \(E^{(M)}\) is controlled by the tail of the semi-discrete kernel and \(\|\!E^{(M)}\!\|_2\) decays with \(M-N\) according to the kernel decay exponent \(r_\alpha=\min(2,1+\alpha)\) [2605.16749].

Pseudo-differential operators supply a broader Fourier-structured class. Generic PDOs can be block-encoded via QFT, phase multiplication, and arithmetic evaluation of the symbol \(a(x,\xi)\), but this yields normalization \(2^{pd/2}C_a\) and exponentially small success probability. Separable symbols \(a(x,\xi)=\alpha(x)\beta(\xi)\) reduce the normalization to \(C_\alpha C_\beta\) and achieve \(\Theta(1)\) success probability, while dimension-wise fully separable symbols admit explicit QET constructions with \(O(d)\) ancillas and gate complexity \(O(p\sum \deg + p^2d)\) [2301.08908].

## 4. Structured matrices, tensor networks, and compressed linear algebra

Rank-structured matrix classes give rise to distinct block-encoding input models. For one-pair semiseparable matrices \(S(u,v)\), an exact factorization
\[
S(u,v)=D_u\,L\,\Delta_z\,L^T\,D_u
\]
supports a block-encoding assembled from diagonal, inverse-diagonal, lower-triangular, and difference-diagonal pieces. The final semiseparable construction uses \(2\log_2N+7\) ancillas, has polylogarithmic depth, and normalization
\[
\alpha=\frac{2N^2M_u^2M_v}{c^2m_u},
\]
with additive spectral-norm error controlled by fixed-point and arcsin approximation errors [2603.19130].

Matrix product operators define another major input model. One compiler dilates each MPO tensor into a \((D+2)\)-qubit unitary, with \(D=\lceil\log\chi\rceil\) determined by the bond dimension \(\chi\). The full chain uses \(L+D\) ancillas and \(O(L\chi^2)\) one- and two-qubit gates, while the global normalization is the product of per-tensor normalizations \(N_{\mathrm{MPO}}=\prod_{\ell=1}^L N_\ell\). The block is exact inside the designated ancilla subspace, although the postselection success probability typically decays exponentially with \(L\) because \(N_{\mathrm{MPO}}\) grows multiplicatively [2312.08861].

A more recent MPO perspective treats tensor networks as compressed virtual-path LCU programs. Expanding each local MPO tensor into a unitary operator basis induces a path sum
\[
M=\sum_{\gamma\in\Gamma} c_\gamma P_\gamma,
\]
with path normalization \(\alpha_{\mathrm{MPO}}=\sum_{\gamma\in\Gamma}|c_\gamma|\). Conditional PREP and SELECT stages can then be compiled directly from the parent MPO, with cost \(O(Nq\chi^2)\) rather than explicit \(O(N^K)\) Pauli-string growth for a degree-\(K\) polynomial expansion, provided the bond dimension and path normalization remain mild [2606.19083].

For dense classical matrices without sparsity or low-rank structure, BITBLE organizes state-preparation unitaries in binary trees and decouples multiplexors by Walsh–Hadamard/Gray-code linear algebra. Its exact normalization can be either \(\|A\|_F\) or \(\mu_p(A)\), the classical preprocessing time is \(O(n2^{2n})\) with memory \(\Theta(2^{2n})\), and the ancilla count is only \(a=n\) or \(a=n+2\) depending on the variant [2504.05624].

## 5. Many-body operator models and direct algebraic constructions

In second quantization, the input model is often built around the algebra of creation and annihilation operators rather than around matrix entries. One recent construction for general second-quantized Hamiltonians combines a SWAP-based sparsity oracle \(O_C\) with SELECT-SWAP data lookup for \(O_A\), giving per-oracle T-count \(\tilde O(n+\sqrt{L})\). The same framework targets the \(\eta\)-particle sector directly through an occupation-detection oracle \(O_{\mathrm{occ}}\), reducing the subnormalization from \(O(L)\) to \(O(n^2\eta^2)\) for general one- and two-body Hamiltonians, with corresponding reductions to \(O(n\eta)\) for one-body terms and \(O(\eta^2)\) for number-operator products [2510.08644].

LOBE block-encodes fermionic and bosonic ladder operators directly, avoiding Pauli-basis expansion. In that framework, single fermionic ladder operators and fermionic products have \(\lambda=1\), while bosonic single-mode ladder operators have \(\lambda=\sqrt{\Omega}\) and products \((a^\dagger)^R a^S\) have \(\lambda=\Omega^{(R+S)/2}\), where \(\Omega\) is the bosonic cutoff. The reported T-counts scale linearly with locality and as \(O(\log\Omega)\) in the bosonic register width, while benchmarks on quartic oscillator, \(\phi^4\), and Yukawa models show fewer non-Clifford gates, fewer ancillas, and lower rescaling factors than Pauli-expansion approaches [2503.11641].

Bosonic lattice Hamiltonians can also be block-encoded by signal-processing methods themselves. QSVT-based encodings use diagonal block-encodings of \(\hat\varphi\), \(\hat\pi^{(D)}\), and \(\hat\varphi_i-\hat\varphi_j\) as primitives; QETU-based constructions work instead from controlled exponentials; LOVE-LCU realizes diagonal functions exactly via
\[
\frac{f(\hat\xi)}{\beta}
=
\frac{e^{i\arccos(f(\hat\xi)/\beta)}+e^{-i\arccos(f(\hat\xi)/\beta)}}{2}.
\]
The reported conclusion is that QSVT has the best asymptotic scaling in qubits per site, whereas LOVE-LCU outperforms the alternatives for operators acting on up to \(\lesssim 11\) qubits [2408.16824].

Linear combinations of Pauli strings admit yet another algebraic input model. A stabilizer-based construction first transforms the Pauli strings into a pairwise anti-commuting set, making the normalized linear combination unitary, and then uses a correction transformation on an ancilla register to restore the original strings. The ancilla requirement scales logarithmically with the number of Pauli terms in the basic version, and larger ancilla registers can reduce circuit complexity further [2601.05740].

## 6. Resource trade-offs, ambiguities, and current directions

Recent work makes clear that “the” block-encoding input model is not a single model but a family of access assumptions with different bottlenecks. One recurrent ambiguity is the relationship between subnormalization and practical cost. Smaller \(\alpha\) is algorithmically advantageous, but it does not by itself imply a cheaper circuit. Single-ancilla exact dense-matrix block-encoding based on diagonal matrix migration attains spectral-norm subnormalization and a leading C-NOT count \((11/48)4^n\); for rank-\(K\) matrices this drops to \((K+11/12)2^n\). These bounds improve on earlier exact synthesis constants, yet they still scale exponentially in \(n\), reflecting the intrinsic cost of dense unstructured inputs [2603.16492].

A second ambiguity concerns postselection versus normalization. Constant subnormalization does not guarantee constant success probability. The DoG construction has \(\lambda=2\) independent of \(N\), \(D\), and \(|T|\), but its exact success probability is power-spectrum weighted and scales as \(\Theta(h^4)\) for smooth inputs on finer periodic grids [2604.09538]. The explicit finite-difference Laplacian encoding exhibits the same \(h^4\) behavior under the stated regularity assumptions [2509.02429]. This shows that success probability may be controlled by the input state and operator spectrum even when the block-encoding normalization is structurally optimal.

A third trade-off is the balance between quantum resources and classical preprocessing. BITBLE uses only a few ancillas but requires \(O(n2^{2n})\) classical preprocessing time and \(\Theta(2^{2n})\) memory [2504.05624]. Dictionary-based sparse block encoding achieves logarithmic circuit depth, but only by spending \(O(n^2s)\) ancillas [2405.18007]. Variational block-encoding can produce exact single-ancilla encodings with parameter counts close to the degrees of freedom of the target matrix, and symmetry-aware ansätze made optimization possible up to \(n=8\) qubits under permutation symmetry; however, the classical optimization itself ceases to be computationally feasible for large system sizes [2507.17658].

Open directions in the literature are correspondingly diverse. Explicit structured encodings invite extensions to nonperiodic boundary conditions through modified shift encodings [2604.09538]. MPO compilers suggest higher-dimensional PEPO generalizations, but that extension is left for future work [2312.08861]. Semiseparable constructions point toward multi-pair semiseparable, HSS, and HODLR variants [2603.19130]. This suggests that the evolution of block-encoding input models is likely to proceed less through a single universal oracle and more through increasingly specialized compilations that preserve the algebraic, geometric, or tensor-network structure of the target operator.

Source: https://www.emergentmind.com/topics/block-encoding-input-model