---
title: Kirkwood-Dirac Quasiprobability Distributions
url: https://www.emergentmind.com/topics/kirkwood-dirac-quasiprobability-distributions-49dd6e82-61d3-4982-959e-44011031dabb
type: topic
---

# Kirkwood-Dirac Quasiprobability Distributions

Searching arXiv for recent and foundational papers on Kirkwood-Dirac quasiprobability distributions.
Searching arXiv: kirkwood-dirac quasiprobability distributions geometry positivity coherence metrology
The Kirkwood–Dirac (KD) quasiprobability distribution is a representation of quantum states relative to the eigenbases of two observables, typically denoted \(A\) and \(B\), that assigns to each pair of basis labels a complex number \(Q_{ij}(\rho)=\langle b_j|a_i\rangle\langle a_i|\rho|b_j\rangle\). It is normalized, reproduces the Born-rule marginals for both observables, and is informationally complete whenever all overlaps \(\langle a_i|b_j\rangle\) are nonzero. Unlike a classical joint probability distribution, however, the KD distribution can assume negative or nonreal values. These departures from positivity have become central to contemporary analyses of nonclassicality, quantum advantage, weak values, measurement disturbance, metrology, thermodynamics, scrambling, and contextuality [2403.18899].

## 1. Definition, normalization, and reconstruction

Let \(A=\sum_i a_i|a_i\rangle\langle a_i|\) and \(B=\sum_j b_j|b_j\rangle\langle b_j|\) be two nondegenerate observables on a \(d\)-dimensional Hilbert space, with orthonormal eigenbases \(\{|a_i\rangle\}\) and \(\{|b_j\rangle\}\). For a density operator \(\rho\), the KD quasiprobability matrix is defined by
\[
Q_{ij}(\rho)=\langle b_j|a_i\rangle\,\langle a_i|\rho|b_j\rangle
=\mathrm{Tr}\!\bigl(\Pi_j^b\,\Pi_i^a\,\rho\bigr),
\]
with \(\Pi_i^a=|a_i\rangle\langle a_i|\) and \(\Pi_j^b=|b_j\rangle\langle b_j|\) [2306.00086].

This array satisfies three probability-like identities:
\[
\sum_{i,j}Q_{ij}(\rho)=1,\qquad
\sum_jQ_{ij}(\rho)=\langle a_i|\rho|a_i\rangle,\qquad
\sum_iQ_{ij}(\rho)=\langle b_j|\rho|b_j\rangle.
\]
Thus the total weight is normalized and the marginals reproduce the outcome distributions of \(A\) and \(B\) [2405.17557].

When \(\min_{i,j}|\langle a_i|b_j\rangle|>0\), the KD representation is informationally complete. One reconstruction formula is
\[
\rho=\sum_{i,j}Q_{ij}(\rho)\,\frac{|a_i\rangle\langle b_j|}{\langle b_j|a_i\rangle},
\]
and more generally the KD symbol of an operator \(F\) is
\[
Q_{ij}(F)=\langle a_i|F|b_j\rangle\,\langle b_j|a_i\rangle.
\]
This makes the KD formalism not only a state representation but also a symbol calculus for operators [2306.00086].

A common misconception is that KD quasiprobabilities are merely classical joint distributions written in unusual notation. They are not: the defining ordered product of projectors is generally non-Hermitian, and this is precisely why the entries can be negative or complex.

## 2. Classicality, nonclassicality, and the meaning of KD positivity

A state is called KD-positive, or KD-classical, with respect to the chosen bases if all entries are real and nonnegative:
\[
Q_{ij}(\rho)\ge 0\quad\forall i,j.
\]
In that case the KD distribution is an honest joint probability distribution on the product label set [2405.17557].

In general, \(Q_{ij}(\rho)\) may be negative or nonreal. Several papers in the recent literature connect these features to nonclassicality. Negative or nonreal KD entries have been linked to quantum advantage in weak measurements and contextuality, enhanced sensitivity in quantum metrology, signatures of quantum chaos, scrambling, and thermodynamic irreversibility [2306.00086]. The real part can be negative even when one passes to the Margenau–Hill distribution, while the imaginary part is associated with measurement disturbance and thermodynamic nonclassicality [2009.04468].

For pure states, one exact criterion for KD classicality is phase factorization. Writing
\[
\langle a_i|\psi\rangle=A_i e^{i\alpha_i},\qquad
\langle \psi|b_j\rangle=B_j e^{i\beta_j},\qquad
U^{AB}_{ij}=\langle a_i|b_j\rangle=|U_{ij}|e^{i\theta_{ij}},
\]
a pure state \(|\psi\rangle\) is KD classical if and only if there exist phases \(\{\alpha_i\}\) and \(\{\beta_j\}\) such that
\[
\theta_{ij}\equiv \alpha_i+\beta_j\pmod{2\pi}
\]
for every relevant pair \((i,j)\) on the supports of the amplitudes [2210.02876]. This criterion formalizes the statement that positivity requires global phase compatibility across all nonzero KD entries.

The literature also clarifies that noncommutation alone does not suffice for KD nonclassicality. Pairwise noncommutation of \(\rho\), \(A\), and \(B\) can occur while the KD distribution remains classical, and tighter counting conditions are required to force negativity or nonreality [2009.04468]. This corrects an older intuition that any incompatibility of observables automatically implies nonpositive KD values.

## 3. Geometry of KD-positive states

The set of KD-positive states,
\[
\mathcal C_{A,B}=\{\rho:Q_{ij}(\rho)\ge 0\ \forall i,j\},
\]
is a convex subset of the density-operator space. Its geometry has become a central theme because it marks the boundary between states admitting a classical-probability interpretation in a chosen KD representation and those exhibiting KD nonpositivity [2306.00086].

A major result is that in several regimes this set is exactly the convex hull of the eigenprojectors of the two reference bases:
\[
\mathcal C_{A,B}=\mathrm{conv}(\mathcal A\cup\mathcal B),
\]
where \(\mathcal A=\{|a_i\rangle\langle a_i|\}\) and \(\mathcal B=\{|b_j\rangle\langle b_j|\}\). This equality holds under the weak-incompatibility condition
\[
m_{A,B}=\min_{i,j}|\langle a_i|b_j\rangle|>0
\]
in the following cases: \(d=2\); an open and dense set of bases in \(d=3\); discrete-Fourier-transform bases in prime dimension; and any pair of bases sufficiently near one of these cases [2306.00086].

For random pairs of orthonormal bases chosen independently according to Haar measure, the situation is even sharper: with probability one, the KD-positive set is the convex hull of the \(2d\) basis projectors. In that case the free-state set is a \((2d-1)\)-dimensional simple polytope with exactly \(2d\) vertices [2405.17557]. This shows that, generically, KD classicality is “minimal” in the sense that no additional extreme states appear beyond the obvious basis eigenstates.

The following summary captures the principal geometric regimes described in the literature.

| Regime | Characterization of \(\mathcal C_{A,B}\) | Source |
|---|---|---|
| \(d=2\) with \(m_{A,B}>0\) | \(\mathrm{conv}(\mathcal A\cup\mathcal B)\) | [2306.00086] |
| \(d=3\), open dense set of bases | \(\mathrm{conv}(\mathcal A\cup\mathcal B)\) | [2306.00086] |
| Prime \(d\), DFT overlap matrix | \(\mathrm{conv}(\mathcal A\cup\mathcal B)\) | [2306.00086] |
| Haar-random bases | With probability one, convex hull of the \(2d\) eigenprojectors | [2405.17557] |

This geometry has direct resource-theoretic significance. If the KD-positive set is just a small polytope, then almost every state lies outside it; the corresponding KD negativity is then generic rather than exceptional [2405.17557].

## 4. Pure-state structure, Fourier pairs, and mixed extremal states

The pure-state sector has a particularly rich classification. For KD-classical pure states, support sizes in the two bases play a decisive role. If
\[
n_A(\psi)=|\{i:\langle a_i|\psi\rangle\neq 0\}|,\qquad
n_B(\psi)=|\{j:\langle b_j|\psi\rangle\neq 0\}|,
\]
then a general block-structure theorem constrains the allowed supports and recovers the support-uncertainty bound \(n_A(\psi)+n_B(\psi)\le d+s\), with \(s\le d/2\), after a suitable decomposition of the relevant submatrix of the transition matrix \(U^{AB}\) into nonoverlapping nonnegative blocks [2210.02876].

For mutually unbiased bases, especially discrete-Fourier-transform pairs,
\[
\langle a_i|b_j\rangle=\frac{1}{\sqrt d}e^{2\pi i\,ij/d},
\]
the classification simplifies dramatically. In that setting a pure state is KD classical if and only if
\[
n_A(\psi)\,n_B(\psi)=d.
\]
This resolves the conjecture of De Bièvre et al. for DFT bases in arbitrary dimension [2210.02876].

A related later result studies DFT-based KD classicality through a directed-graph construction. In dimension \(d=p^r\), the full set of KD-classical states equals the convex hull of KD-classical pure states. More generally, for arbitrary \(d\), along any path in the directed graph associated with a factorization structure of \(d\), the intersection of the KD-classical state set with the real span of the path-associated KD-classical pure states is exactly the convex hull of those pure states [2603.13863]. This suggests a refined combinatorial organization of KD-classical sectors in composite dimensions.

Not all dimensions and basis pairs are so rigid. When the overlap matrix \(U\) is real orthogonal in dimension \(d\ge 3\), the real-linear space of operators with real KD symbol has dimension \(d(d+1)/2\), which is strictly larger than \(2d-1\). In this regime one can have
\[
\mathrm{PurePos}\subsetneq \mathrm{Ext}(\mathcal C_{A,B})\subsetneq \mathcal C_{A,B},
\]
so mixed extreme KD-positive states must exist [2306.00086].

The explicit spin-1 example is particularly instructive. Taking \(\mathcal A\) as the \(J_z\) eigenbasis and \(\mathcal B\) as the eigenbasis of a rotated spin component \(n\!\cdot\!J\) about \(n=(1,1,1)/\sqrt 3\) by angle \(\pi\), the real orthogonal overlap matrix is
\[
U^*=\frac13
\begin{pmatrix}
-1&2&2\\
2&-1&2\\
2&2&-1
\end{pmatrix}.
\]
In this case there exist KD-positive mixed states of the form
\[
\rho(x)=\frac13 I+xF_\perp
\]
that remain KD-positive for an open interval of \(x>0\) yet lie outside \(\mathrm{conv}(\mathcal A\cup\mathcal B)\). Hence they are mixed extreme points that are not convex combinations of pure KD-positive states [2306.00086].

This phenomenon matters conceptually because it distinguishes “KD-positive” from “mixtures of pure KD-positive states.” The two notions coincide generically, but not universally.

## 5. Quantifying nonpositivity, coherence, and related witnesses

A widely used scalar quantifier is the total KD nonpositivity
\[
\mathcal N(Q)=\sum_{i,j}|Q_{ij}|,
\]
or equivalently \(-1+\sum_{i,j}|Q_{ij}|\) depending on convention. It equals its classical minimum exactly when the KD distribution is real and nonnegative [2502.11784; 2009.04468]. For \(k\) projective measurements, the maximum total nonclassicality is
\[
\max \mathcal N=d^{(k-1)/2}-1,
\]
achieved if and only if each pair of successive bases is mutually unbiased and the state is pure with equal overlap \(1/\sqrt d\) with every vector in the first and last bases [2009.04468].

Recent work has developed convex-roof constructions to extend pure-state witnesses to mixed states. The convex roof of support uncertainty,
\[
\widetilde U_s(\rho)=\min_{\rho=\sum_k p_k|\psi_k\rangle\langle\psi_k|}\sum_k p_k\,U_s(\psi_k),
\]
provides a nonfaithful witness for the convex hull of pure KD-positive states, while the convex roof of the total KD nonpositivity,
\[
\widetilde N(\rho)=\min_{\rho=\sum_k p_k|\psi_k\rangle\langle\psi_k|}\sum_k p_k\,N(\psi_k),
\]
is faithful for that convex hull: \(\widetilde N(\rho)=1\) if and only if \(\rho\) lies in the convex hull of pure KD-positive states [2407.04558].

Moment-based detection criteria provide another route. Defining raw moments
\[
q_n=\sum_{i,j}[Q_{ij}(\rho)]^n,
\]
classical positivity implies
\[
(q_2)^2\le q_3.
\]
More generally, the Hankel matrices \(H_m(q)\) built from the moments must be positive semidefinite if the KD table is pointwise nonnegative, so \(\det H_m(q)<0\) certifies KD nonpositivity [2506.08107]. Explicit examples show that higher-order determinants can detect nonpositivity even when the second-versus-third-order test fails.

KD nonclassicality has also been turned into a coherence monotone. For a fixed incoherent basis \(\{\Pi_j\}\), the KD-nonclassicality coherence is defined by
\[
C_{KD}^{NCl}(\rho;\{\Pi_j\})
=\sup_{\{|b\rangle\}}\left[\sum_{j,b}|\langle b|\Pi_j\rho|b\rangle|-1\right].
\]
It obeys
\[
C_{KD}^{NCl}(\rho;\{\Pi_j\})\le \sum_j\sqrt{p_j}-1=\tfrac12 S_{1/2}(\{p_j\})\le \sqrt d-1,
\]
where \(p_j=\mathrm{Tr}(\Pi_j\rho)\), and for pure states the upper bound is saturated:
\[
C_{KD}^{NCl}(|\psi\rangle;\{\Pi_j\})=\sum_j\sqrt{p_j}-1.
\]
The same framework yields uncertainty-based lower bounds via the Maassen–Uffink relation and trade-off relations between KD-nonclassicality coherences in different bases [2309.09162].

## 6. Operational roles in measurement, tomography, metrology, thermodynamics, and scrambling

The KD distribution has become a unifying operational language across several quantum-information subfields. In measurement theory, the imaginary part of the KD distribution governs disturbance. Under the unitary \(\hat U(\theta)=e^{-i\hat A\theta}\), the rate of change of a later measurement probability is
\[
\partial_\theta\langle f_j|\rho_\theta|f_j\rangle
=2\sum_i a_i\,\Im\,Q_{ij}(\rho_\theta),
\]
so nonreal conditional KD values are directly tied to measurement back-action [2403.18899].

In weak measurement theory, the weak value of an observable is a conditional KD average. Anomalous weak values require KD nonpositivity, and “strange” weak values imply contextuality in hidden-variable descriptions [2403.18899; 2309.09162]. Direct experimental access to KD values can be obtained without post-selection using ancilla-assisted cycle-test circuits that estimate the relevant Bargmann invariants. In these schemes, the real and imaginary parts of \(\mathrm{KD}_\rho(a,b)\) are obtained from ancilla expectation values after a controlled cyclic SWAP, and the same toolkit extends to OTOCs and post-selected Fisher information [2302.00705].

In finite-dimensional tomography, KD enjoys several properties emphasized as comparative advantages over many other quasiprobabilities. It vanishes off the grid of actual eigenvalue pairs, obeys correct marginals, satisfies a Cauchy–Schwarz bound
\[
|KD_\rho(a,b)|\le \sqrt{P_\rho^A(a)P_\rho^B(b)},
\]
and in spin-\(\tfrac12\) and spin-1 systems two noncommuting observables suffice to determine the state completely from the KD distribution [2309.06836].

In quantum metrology, KD negativity and imaginarity have been linked to enhanced performance. Standard Fisher-information expressions can be written in terms of conditional KD distributions, and post-selected Fisher information can be expressed as a quasiprobability covariance of an extended KD distribution; KD negativity enables unbounded compressive concentration of Fisher information into selected trials [2403.18899]. The 2020 bounds on achievable KD nonclassicality were explicitly motivated by metrological advantage, including the observation that only the negative-real part can underlie nonclassical Fisher-information boosts, whereas the imaginary part encodes disturbance [2009.04468].

In quantum thermodynamics, KD-based work quasiprobabilities replace the two-point-measurement protocol when initial coherence makes ordinary work statistics inadequate. For a driven Hamiltonian,
\[
q_{if}(\rho)=\mathrm{Tr}\!\bigl[U^\dagger(t)\Pi_f(t)U(t)\Pi_i(0)\rho\bigr],
\qquad
P(W)=\sum_{i,f}q_{if}(\rho)\,\delta\!\bigl(W-(E_f(t)-E_i(0))\bigr).
\]
This quasiprobability reproduces the unperturbed average work and its higher moments while allowing complex values that quantify noncommutativity [2405.21041]. Interferometric experiments in an NV-center electron-nuclear spin platform reconstructed the characteristic function and extracted the quasiprobability distribution of work, including its first and second moments [2405.21041]. Related KDQ constructions have also been developed for coherent collision models, where negativities and imaginary parts diagnose quantum traits in energy exchange and work statistics [2503.07759].

In scrambling and chaos, the out-of-time-ordered correlator can be expressed as a moment of an extended KD distribution. The resulting quasiprobability exhibits negative and nonreal behavior, distinguishes chaotic from integrable dynamics, and develops symmetry-breaking “pitchfork” bifurcations at the onset of scrambling [1704.01971]. This operational role is now one of the standard applications summarized in the broader review literature [2403.18899].

## 7. Relations to other quasiprobabilities and broader conceptual developments

The KD distribution is frequently contrasted with the Wigner function. Unlike the Wigner formalism, which is tied to specific phase-space structures, KD can be defined for arbitrary observables or bases [2403.18899]. For finite-state systems, one argument for its special status is that it stays supported only on the actual spectral grid and can encode the full density operator from as few as two noncommuting observables, with the imaginary part essential for full distinguishability in qubits [2309.06836].

There are also direct structural links between KD and discrete Wigner theory. For KD distributions defined in qudit Fourier bases, the discrete Fourier transform of the KD table yields expectation values of Weyl–Heisenberg operators, and a further symplectic Fourier transform gives the odd-dimensional discrete Wigner distribution. This provides a KD-to-Wigner mapping without reconstructing the density matrix [2502.11784].

Resource-theoretic and simulation-theoretic questions have sharpened these comparisons. In KD representations based on mutually unbiased bases, pure KD-positive states always have uniform distributions on their support of size \(d\), “in full analogy to the stabilizer states” [2502.11784]. At the same time, positivity preservation under unitary evolution need not coincide with stochastic quasiprobability evolution: there exist unitaries that preserve KD positivity of every state while inducing superoperators with negative entries, and classical sampling algorithms can become exponentially inefficient even when the state remains KD positive throughout the evolution [2502.11784]. This suggests that KD positivity alone does not determine classical simulability.

More abstractly, KD has been situated between classical and “postquantum” quasiprobabilities. Under mild assumptions one obtains a strict hierarchy
\[
\mathcal P_{\rm cl}\subsetneqq \mathcal P_{\rm KD}\subsetneqq \mathcal P_{\rm post},
\]
together with bounds such as
\[
\sum_{x,y}|q_{xy}|^2\le 1,\qquad
\sum_{x,y}|q_{xy}|\le \sqrt N,
\]
and extensions of the \(L^2\)-norm bound to arbitrarily many measurements [2504.09238]. These results formalize the statement that KD enlarges classical probability theory, but not arbitrarily.

A further characterization singles out KD among all Born-compatible quasiprobability representations on the joint spectrum of two observables. For such representations, each observable \(\hat X\) has an associated conditional expectation given \(\hat B\); among them, only the left-KD representation has the property that this conditional expectation coincides with the best predictor of \(\hat X\) by a function of \(\hat B\) for all \(\hat X\) [2511.01996]. This identifies a precise statistical-optimality property that is unique to KD.

Taken together, these developments present the KD quasiprobability distribution as a flexible but highly structured framework. Its central mathematical tension is that it retains normalization, correct marginals, and, in many settings, informational completeness, while relaxing positivity and reality. Its central physical tension is that those relaxed features are precisely what permit it to encode disturbance, incompatibility, coherence, contextuality, and operational quantum advantage.

Source: https://www.emergentmind.com/topics/kirkwood-dirac-quasiprobability-distributions-49dd6e82-61d3-4982-959e-44011031dabb