Quantum Block-Encoding Input Model
- Block-Encoding Input Model is a framework that encodes matrix or operator data into a unitary, enabling quantum algorithms like QSVT and qubitization.
- It details various access models—including oracle-based, state-preparation, and arithmetic methods—each with distinct normalization, ancilla, and circuit depth requirements.
- The approach emphasizes resource trade-offs and specialized constructions, highlighting the interplay between classical preprocessing, implementation cost, and success probability.
The block-encoding input model is the mechanism by which a quantum algorithm receives matrix or operator data as a unitary-access primitive. In its standard form, a unitary acting on ancillas and system is an -block-encoding of an operator when ; in the exact case, the top-left block of is . In recent work, the phrase “input model” denotes not only this definition but also the concrete assumptions under which such a unitary is compiled: sparse-access and state-preparation oracles, explicit arithmetic circuits, Fourier-symbol oracles, tensor-network contractions, and variationally trained unitaries all instantiate different block-encoding input models with sharply different normalization factors, ancilla counts, gate depths, and classical preprocessing requirements (Shi et al., 2024, Li et al., 2023, Mahmud et al., 10 Apr 2026).
1. Formal role in quantum algorithms
Block encoding is the standard access primitive for QSVT, qubitization, and related polynomial-transform methods. In this framework, is the subnormalization factor, and the success amplitude in a postselection picture is ; when acts on , the ancilla-zero component is 0. In QSP- and QSVT-based simulation, the number of uses of the block-encoding scales linearly in 1 up to polylogarithmic factors in the target precision, so reducing 2 directly reduces query complexity (Liu et al., 9 Oct 2025).
The same formalism accommodates non-Hermitian inputs. One standard route is Hermitian dilation, 3, which allows QSVT to act on singular values while preserving a Hermitian signal operator. This is used explicitly in quantum linear-algebra settings such as Kalman filtering, where addition, multiplication, and inversion are all expressed through block-encodings and composed inside a unified framework (Shi et al., 2024).
A central distinction in the literature is therefore between the abstract definition of a block-encoding and the concrete way the unitary is realized. Some constructions assume oracle access to entries or sparse structure; others compile the unitary explicitly from problem structure, so that no qRAM, signed amplitude loading, or black-box entry oracle is required. The latter trend is especially pronounced in recent work on differential operators, tensor networks, and many-body Hamiltonians (Mahmud et al., 10 Apr 2026).
2. Oracle-based, state-preparation, and arithmetic access models
A large class of input models is oracle-based. In dense-matrix settings, a common assumption is access to state-preparation oracles 4 and 5 that prepare row- and column-weighted superpositions; their composition yields a 6-block-encoding of 7 with polylogarithmic query time in the matrix dimensions, provided the data are stored in a quantum-accessible structure (Shi et al., 2024). In sparse and second-quantized settings, the input model is instead phrased in terms of a sparsity oracle 8 and an amplitude oracle 9. For second-quantized Hamiltonians, SWAP-based implementations of 0 and SELECT-SWAP data lookup for 1 reduce the T-count per oracle invocation to 2 in the number of interaction terms 3 (Liu et al., 9 Oct 2025).
A different line of work replaces generic oracles by arithmetic descriptions of structure. For matrices with repeated values and patterned sparsity, one can specify nonzero entries by a value label 4, a multiplicity label 5, and reversible arithmetic maps between 6 and row/column coordinates. In that model, the dominant data-loading cost depends on the number 7 of distinct values rather than the matrix dimension, and different schemes produce different subnormalizations: a base scheme with 8, a preamplified scheme with 9, and a PREP/UNPREP scheme with 0 when the requisite commutation conditions hold (Sünderhauf et al., 2023).
The dictionary-based sparse model pushes this idea further. There, a sparse matrix is organized into classes of repeated values 1, together with injective maps 2 that specify where those values occur. The resulting unitary
3
block-encodes 4 with subnormalization 5, depth 6, and ancilla count 7, where 8 is the number of nonzeros and 9 is the number of dictionary items (Yang et al., 2024).
Approximate state-preparation approaches remain relevant for unstructured sparse data. S-FABLE block-encodes 0, then conjugates by outer Hadamards to recover a block-encoding of 1; LS-FABLE avoids the quadratic classical overhead by directly inserting scaled sparse entries into the rotation angles. For unstructured sparse matrices with 2 nonzeros, the reported empirical behavior is approximately 3 rotation gates and 4 CNOT gates after compression (Kuklinski et al., 2024).
| Model | Access assumption | Characteristic feature |
|---|---|---|
| State-preparation oracles (Shi et al., 2024) | 5 for row/column amplitudes | Frobenius-norm normalization |
| Sparse oracle + amplitude oracle (Liu et al., 9 Oct 2025) | 6 | 7 T-count |
| Arithmetic structured matrices (Sünderhauf et al., 2023) | reversible maps 8 | loading depends on distinct values 9 |
| Dictionary sparse (Yang et al., 2024) | value classes and injective maps 0 | 1, depth 2 |
| S-FABLE / LS-FABLE (Kuklinski et al., 2024) | classical angle preprocessing or sparse-entry access | aggressive circuit compression for sparse 3 |
3. Explicit structure-exploiting encodings for differential and Fourier operators
A prominent current direction is to eliminate generic data oracles entirely by exploiting operator structure. For the Difference-of-Gaussian operator on a periodic grid, the coefficients split into two normalized discrete Gaussian distributions 4 and 5. Exact Gaussian state-preparation circuits 6 and 7, a one-qubit branch indicator, a single Pauli-8 gate to encode the minus sign, and controlled cyclic shifts together yield an exact block-encoding of
9
with 0, independent of grid size 1, spatial dimension 2, and stencil width 3. The same work derives an exact success probability
4
and shows 5 for smooth inputs as the periodic grid is refined (Mahmud et al., 10 Apr 2026).
For finite-difference discretizations of the Laplacian on periodic grids, the input model is fully explicit: ancillas prepare fixed superpositions, and controlled cyclic shifts realize the stencil. In one dimension this gives 6; in 7 dimensions the subnormalization is 8, with 9 T-gate complexity and 0 under 1 regularity assumptions (Sturm et al., 2 Sep 2025).
QFT-based models access operators through their Fourier symbols. For bounded-domain fractional Laplacians with open, zero-extension boundary conditions, the native QFT implements a periodic circulant surrogate rather than the Toeplitz truncation. Zero-padding into an 2-point periodic register and compressing back to the physical 3-point subspace produces
4
where the residual 5 is controlled by the tail of the semi-discrete kernel and 6 decays with 7 according to the kernel decay exponent 8 (Javanmard et al., 16 May 2026).
Pseudo-differential operators supply a broader Fourier-structured class. Generic PDOs can be block-encoded via QFT, phase multiplication, and arithmetic evaluation of the symbol 9, but this yields normalization 0 and exponentially small success probability. Separable symbols 1 reduce the normalization to 2 and achieve 3 success probability, while dimension-wise fully separable symbols admit explicit QET constructions with 4 ancillas and gate complexity 5 (Li et al., 2023).
4. Structured matrices, tensor networks, and compressed linear algebra
Rank-structured matrix classes give rise to distinct block-encoding input models. For one-pair semiseparable matrices 6, an exact factorization
7
supports a block-encoding assembled from diagonal, inverse-diagonal, lower-triangular, and difference-diagonal pieces. The final semiseparable construction uses 8 ancillas, has polylogarithmic depth, and normalization
9
with additive spectral-norm error controlled by fixed-point and arcsin approximation errors (Antonioli et al., 19 Mar 2026).
Matrix product operators define another major input model. One compiler dilates each MPO tensor into a 0-qubit unitary, with 1 determined by the bond dimension 2. The full chain uses 3 ancillas and 4 one- and two-qubit gates, while the global normalization is the product of per-tensor normalizations 5. The block is exact inside the designated ancilla subspace, although the postselection success probability typically decays exponentially with 6 because 7 grows multiplicatively (Nibbi et al., 2023).
A more recent MPO perspective treats tensor networks as compressed virtual-path LCU programs. Expanding each local MPO tensor into a unitary operator basis induces a path sum
8
with path normalization 9. Conditional PREP and SELECT stages can then be compiled directly from the parent MPO, with cost 00 rather than explicit 01 Pauli-string growth for a degree-02 polynomial expansion, provided the bond dimension and path normalization remain mild (Dumitrescu, 17 Jun 2026).
For dense classical matrices without sparsity or low-rank structure, BITBLE organizes state-preparation unitaries in binary trees and decouples multiplexors by Walsh–Hadamard/Gray-code linear algebra. Its exact normalization can be either 03 or 04, the classical preprocessing time is 05 with memory 06, and the ancilla count is only 07 or 08 depending on the variant (Li et al., 8 Apr 2025).
5. Many-body operator models and direct algebraic constructions
In second quantization, the input model is often built around the algebra of creation and annihilation operators rather than around matrix entries. One recent construction for general second-quantized Hamiltonians combines a SWAP-based sparsity oracle 09 with SELECT-SWAP data lookup for 10, giving per-oracle T-count 11. The same framework targets the 12-particle sector directly through an occupation-detection oracle 13, reducing the subnormalization from 14 to 15 for general one- and two-body Hamiltonians, with corresponding reductions to 16 for one-body terms and 17 for number-operator products (Liu et al., 9 Oct 2025).
LOBE block-encodes fermionic and bosonic ladder operators directly, avoiding Pauli-basis expansion. In that framework, single fermionic ladder operators and fermionic products have 18, while bosonic single-mode ladder operators have 19 and products 20 have 21, where 22 is the bosonic cutoff. The reported T-counts scale linearly with locality and as 23 in the bosonic register width, while benchmarks on quartic oscillator, 24, and Yukawa models show fewer non-Clifford gates, fewer ancillas, and lower rescaling factors than Pauli-expansion approaches (Simon et al., 14 Mar 2025).
Bosonic lattice Hamiltonians can also be block-encoded by signal-processing methods themselves. QSVT-based encodings use diagonal block-encodings of 25, 26, and 27 as primitives; QETU-based constructions work instead from controlled exponentials; LOVE-LCU realizes diagonal functions exactly via
28
The reported conclusion is that QSVT has the best asymptotic scaling in qubits per site, whereas LOVE-LCU outperforms the alternatives for operators acting on up to 29 qubits (Kane et al., 2024).
Linear combinations of Pauli strings admit yet another algebraic input model. A stabilizer-based construction first transforms the Pauli strings into a pairwise anti-commuting set, making the normalized linear combination unitary, and then uses a correction transformation on an ancilla register to restore the original strings. The ancilla requirement scales logarithmically with the number of Pauli terms in the basic version, and larger ancilla registers can reduce circuit complexity further (Schillo et al., 9 Jan 2026).
6. Resource trade-offs, ambiguities, and current directions
Recent work makes clear that “the” block-encoding input model is not a single model but a family of access assumptions with different bottlenecks. One recurrent ambiguity is the relationship between subnormalization and practical cost. Smaller 30 is algorithmically advantageous, but it does not by itself imply a cheaper circuit. Single-ancilla exact dense-matrix block-encoding based on diagonal matrix migration attains spectral-norm subnormalization and a leading C-NOT count 31; for rank-32 matrices this drops to 33. These bounds improve on earlier exact synthesis constants, yet they still scale exponentially in 34, reflecting the intrinsic cost of dense unstructured inputs (Li et al., 17 Mar 2026).
A second ambiguity concerns postselection versus normalization. Constant subnormalization does not guarantee constant success probability. The DoG construction has 35 independent of 36, 37, and 38, but its exact success probability is power-spectrum weighted and scales as 39 for smooth inputs on finer periodic grids (Mahmud et al., 10 Apr 2026). The explicit finite-difference Laplacian encoding exhibits the same 40 behavior under the stated regularity assumptions (Sturm et al., 2 Sep 2025). This shows that success probability may be controlled by the input state and operator spectrum even when the block-encoding normalization is structurally optimal.
A third trade-off is the balance between quantum resources and classical preprocessing. BITBLE uses only a few ancillas but requires 41 classical preprocessing time and 42 memory (Li et al., 8 Apr 2025). Dictionary-based sparse block encoding achieves logarithmic circuit depth, but only by spending 43 ancillas (Yang et al., 2024). Variational block-encoding can produce exact single-ancilla encodings with parameter counts close to the degrees of freedom of the target matrix, and symmetry-aware ansätze made optimization possible up to 44 qubits under permutation symmetry; however, the classical optimization itself ceases to be computationally feasible for large system sizes (Rullkötter et al., 23 Jul 2025).
Open directions in the literature are correspondingly diverse. Explicit structured encodings invite extensions to nonperiodic boundary conditions through modified shift encodings (Mahmud et al., 10 Apr 2026). MPO compilers suggest higher-dimensional PEPO generalizations, but that extension is left for future work (Nibbi et al., 2023). Semiseparable constructions point toward multi-pair semiseparable, HSS, and HODLR variants (Antonioli et al., 19 Mar 2026). This suggests that the evolution of block-encoding input models is likely to proceed less through a single universal oracle and more through increasingly specialized compilations that preserve the algebraic, geometric, or tensor-network structure of the target operator.