- The paper systematically compares non-Clifford resource costs for qudit and qubit encodings of diagonal quadratic evolutions using product-formula and LCU constructions.
- Product-formula methods offer no scalable qudit advantage because qudit synthesis would need an exponentially smaller per-rotation cost, although dimensions d=3 and d=5 can provide constant-factor savings.
- LCU methods favor qubits asymptotically, but finite-dimension qudit benefits appear for physically relevant cutoffs, reaching roughly d≤19 without switching and d≤21 with idealized code conversion.
This paper presents a systematic fault-tolerant resource comparison between single logical d-level qudit encodings and registers of nb=⌈log2d⌉ logical qubits for implementing onsite diagonal quadratic evolutions, exemplified by U=e−itϕx2 in a uniform field-amplitude (Jordan–Lee–Preskill) digitization of a real scalar field (2604.26792). The resource metric throughout is the non-Clifford gate count after synthesis into a discrete fault-tolerant gate set, and the analysis is carried out in two simulation regimes: product-formula (Trotter) steps and LCU/block-encoding constructions. The central finding is deliberately qualified: within the constructive models studied, qudits do not provide an asymptotic advantage for this operator class, but explicit finite-d break-even conditions identify low-dimensional windows in which constant-factor savings are possible.
Problem setup
The digitized field operator ϕx acts diagonally on a symmetric grid of d=2M+1 levels with eigenvalues λn=−ϕmax+nδϕ. The target unitary e−itϕx2 is representative of a broader class of onsite quadratic diagonal evolutions e−itg(N) arising in lattice Hamiltonians, gauge theories, condensed matter, and quantum chemistry. Qubit implementations use binary or signed-binary embeddings on nb qubits; qudit implementations act natively on the nb=⌈log2d⌉0-level truncation. Because tight synthesis bounds for general single-qudit rotations are not known uniformly in nb=⌈log2d⌉1, the authors express all qudit costs in terms of embedded two-level nb=⌈log2d⌉2 rotations nb=⌈log2d⌉3 and parameterize their synthesis cost as nb=⌈log2d⌉4, against the qubit baseline nb=⌈log2d⌉5. The resulting thresholds are framed explicitly as compiler targets rather than closed-form costs, and the finite-nb=⌈log2d⌉6 crossover analyses restrict to prime local dimensions, reflecting the state of the qudit Clifford+nb=⌈log2d⌉7/magic-state-distillation literature.
On the qubit side, a Trotter step of nb=⌈log2d⌉8 requires at most nb=⌈log2d⌉9 synthesized U=e−itϕx20 gates: squaring the binary expansion of U=e−itϕx21 produces only one- and two-body U=e−itϕx22 and U=e−itϕx23 Pauli terms, each contributing one rotation. On the qudit side, any diagonal unitary decomposes into exactly U=e−itϕx24 adjacent two-level U=e−itϕx25 rotations via a Givens-rotation argument, with unique angles modulo U=e−itϕx26; for the quadratic phase these angles are generically all nontrivial.
The asymptotic consequence is stark. For the qudit construction to beat the qubit baseline asymptotically, the per-primitive synthesis cost of embedded U=e−itϕx27 rotations would need to be smaller than qubit-optimal U=e−itϕx28 synthesis by a factor U=e−itϕx29 — an exponentially strong advantage in d0. Since volume-counting arguments impose d1 lower bounds and known number-theoretic synthesis procedures achieve matching d2 scaling, the authors conclude this exponential per-rotation advantage is incompatible with the standard synthesis landscape. Phase kickback does not evade the issue, since a diagonal phase function must still be synthesized after coherently computing d3.
The finite-d4 crossover analysis nonetheless identifies a narrow window. At representative precision d5, the break-even prefactor satisfies d6 (d7), d8 (d9), and drops to ϕx0 (ϕx1). Only for ϕx2 and ϕx3 does the threshold exceed the effective qubit-ϕx4 prefactor at matched primitive precision, meaning Regime 1 offers at most a low-ϕx5 opportunity rather than scalable savings.
LCU/block-encoding regime
In Regime 2 the motivation for qudits is structurally stronger: the entangling part of SELECT, ϕx6, is a generalized controlled-ϕx7 and hence Clifford on qudits, whereas incrementers and qudit Fourier transforms are also Clifford. The paper develops two comparisons.
Qubit baseline. Using the signed-binary projector-LCU construction of Su et al. and Spagnoli et al., ϕx8 admits an exact block encoding with normalization ϕx9 and per-call d=2M+10 count d=2M+11, where d=2M+12. This cost grows only as d=2M+13.
Qudit construction. The generalized-Pauli expansion d=2M+14 has irreducible length d=2M+15: all Fourier coefficients d=2M+16 are nonzero for odd d=2M+17, so PREP amounts to generic d=2M+18-dimensional state preparation requiring d=2M+19 embedded two-level rotations. Under the idealized code-switching model — where the λn=−ϕmax+nδϕ0-qubit index register converts freely to a single λn=−ϕmax+nδϕ1-level qudit for the controlled-λn=−ϕmax+nδϕ2 interaction — the qudit side costs λn=−ϕmax+nδϕ3 λn=−ϕmax+nδϕ4 gates per call due to the λn=−ϕmax+nδϕ5-synthesis of the PREP cascade, versus λn=−ϕmax+nδϕ6 for the qubit baseline. Theorem 3 establishes that the ratio of qubit to qudit total cost tends to zero as λn=−ϕmax+nδϕ7: the qubit encoding is asymptotically cheaper, driven entirely by per-call cost growth since both query counts are λn=−ϕmax+nδϕ8 in λn=−ϕmax+nδϕ9 at fixed e−itϕx20 and e−itϕx21.
Finite-d results and code-switching budgets
Weighting per-call costs by the qubitization query count e−itϕx22 yields end-to-end comparisons at precision-dominated (e−itϕx23) and time-dominated (e−itϕx24) points, with e−itϕx25 and e−itϕx26.
In the fixed-encoding model, the condition that the break-even prefactor exceeds the same-precision qubit-e−itϕx27 reference holds only for e−itϕx28 (e−itϕx29) and e−itg(N)0 (e−itg(N)1) in the precision-dominated regime, but extends to all prime dimensions up to e−itg(N)2 in the time-dominated regime, with the largest value e−itg(N)3 against a reference of e−itg(N)4. For context, existing qutrit Clifford+e−itg(N)5 synthesis reports effective prefactors around e−itg(N)6–e−itg(N)7, suggesting the favorable e−itg(N)8 window may be practically reachable, though a direct comparison requires a gate-set conversion not generally available.
Under the code-switching model, the qudit construction wins only for e−itg(N)9 and nb0 in the precision-dominated regime (crossover between nb1 and nb2, with nb3 and absolute savings of roughly nb4 nb5 gates), while in the time-dominated regime it wins for nb6, with the largest relative advantage nb7 and largest absolute savings nb8 nb9 gates at nb=⌈log2d⌉00. These savings translate into per-switch overhead budgets: assuming nb=⌈log2d⌉01 directional switches per query, qudits remain advantageous provided each switch costs fewer than nb=⌈log2d⌉02 non-Cliffords, yielding budgets up to nb=⌈log2d⌉03 nb=⌈log2d⌉04 gates near the most favorable dimensions (e.g., nb=⌈log2d⌉05 at nb=⌈log2d⌉06).
Physical relevance of the favorable windows
The identified windows overlap meaningfully with physically motivated truncations. Klco and Savage estimate nb=⌈log2d⌉07–nb=⌈log2d⌉08 qubits per site before digitization error falls below typical Trotter error; only the lower end (nb=⌈log2d⌉09–nb=⌈log2d⌉10) overlaps the time-dominated break-even region, while nb=⌈log2d⌉11 lies outside it in both models. By contrast, gauge-theory applications often require explicitly small cutoffs: nb=⌈log2d⌉12D scalar QED appears sufficient at nb=⌈log2d⌉13, pure nb=⌈log2d⌉14 gauge theory in nb=⌈log2d⌉15D achieves per-mille plaquette accuracy at nb=⌈log2d⌉16, and trapped-ion qudit simulations of 2D lattice QED operate at nb=⌈log2d⌉17 and nb=⌈log2d⌉18. These cases fall squarely within the favorable regions, motivating fault-tolerant qudit investigations at small prime dimensions.
Limitations and open questions
Several caveats bound the scope of these conclusions. First, the absence of tight, gate-set-dependent synthesis bounds for general single-qudit diagonal operations forces the parameterization by the prefactor nb=⌈log2d⌉19; deriving closed-form costs for arbitrary odd nb=⌈log2d⌉20, or proving whether nb=⌈log2d⌉21 saturates worst-case lower bounds, remains open and is related to NP-complete nearest-unitary decision problems. Second, the Regime 2 comparisons are construction-specific: they compare the native qudit LCU against one standard signed-binary qubit projector-LCU baseline, not optimal lower bounds over all qubit encodings, so alternative qubit constructions could shift the comparison in either direction. Third, the code-switching model assumes free isometric conversion between the index register and the qudit, with invalid computational states unpopulated; the actual fault-tolerant overhead of qubit–qudit code conversion is architecture-dependent and unresolved, though qubit fusion/qudit fission via non-stabilizer resource states offers one possible route whose distillation costs remain unquantified. Fourth, the query-count proxy introduces nb=⌈log2d⌉22 shifts near break-even, and the restriction to prime nb=⌈log2d⌉23 reflects the current synthesis literature rather than a fundamental constraint.
Conclusion
The paper reframes the qudit-versus-qubit question for quadratic diagonal evolutions away from asymptotic claims and toward explicit, finite-nb=⌈log2d⌉24 compiler targets. Within the studied constructions, product-formula implementations would require an exponentially strong per-primitive synthesis advantage for qudits to win, and the LCU setting favors qubits asymptotically in nb=⌈log2d⌉25; nevertheless, meaningful constant-factor savings arise at low dimensions — up to nb=⌈log2d⌉26 in the fixed-encoding time-dominated regime and nb=⌈log2d⌉27 under idealized code-switching — precisely where several lattice gauge theory applications require modest truncations. A definitive verdict awaits dimension-specific synthesis results for embedded nb=⌈log2d⌉28 primitives and architecture-level accounting of code conversion, injection, and distillation overheads (2604.26792).