Conditional Vector Quantile Regression
- Conditional Vector Quantile Regression is a framework that extends scalar quantile regression to multivariate settings using monotone transport maps derived from convex potentials.
- It integrates optimal transport theory to construct a unique, joint quantile function that represents the full conditional distribution while ensuring global monotonicity.
- CVQR has wide applications, including econometrics and manifold learning, and offers robust estimation even under misspecified models through convex-envelope approximations.
Searching arXiv for recent and foundational work on conditional vector quantile regression to ground the article. Conditional Vector Quantile Regression (CVQR) is a conditional multivariate quantile framework that extends scalar quantile regression to vector-valued outcomes by representing conditional distributions through monotone transport maps from a reference rank law to the law of the response given covariates. In the foundational formulation, the central object is the conditional vector quantile function (CVQF), a map that is monotone in the multivariate sense of being the gradient of a convex function in , pushes a non-atomic reference distribution onto the conditional law of , and yields the strong representation almost surely for an appropriate latent rank vector (Carlier et al., 2014). The term “CVQR” is not the authors’ formal label in the foundational paper; the formal regression model is “vector quantile regression” (VQR), which operationalizes the conditional framework through linear or sieve-type specifications of the CVQF (Carlier et al., 2014).
1. Definition and geometric structure
Let and be random vectors, with conditional distribution . Fix a reference, non-atomic distribution on 0 with density 1 and convex support, for example 2 or 3. A conditional vector quantile function is a measurable map
4
such that for each 5, the map 6 is monotone in the sense of being the gradient of a convex function,
7
for some convex potential 8. Equivalently, for all 9 in the support of 0,
1
With 2, the pushforward condition is
3
and the strong representation is
4
Under the non-atomicity and density condition on 5, the CVQF exists and is unique 6-almost everywhere for each 7, as a conditional Brenier map (Carlier et al., 2014).
When each conditional law 8 admits a density, there is also a conditional inverse or rank map,
9
where 0 is the Legendre transform of 1. It satisfies
2
and
3
This gives CVQR a precise multivariate rank structure: 4 is simultaneously a transport coordinate, a conditional rank, and a latent factor with prescribed marginals (Carlier et al., 2014).
A common misconception is that multivariate quantiles can be defined componentwise without loss. The CVQR framework rejects that simplification: dependence across outcome components is encoded through the joint rank 5 and the convex potential 6, not through separate scalar quantiles. In the scalar case 7, the construction collapses to the usual conditional quantile function, but for 8 the monotone map is intrinsically joint rather than coordinatewise (Carlier et al., 2014).
2. Optimal transport formulation
A defining feature of CVQR is that it embeds Monge–Kantorovich optimal transport at its core. Under finite second moments, the CVQF arises as the solution to a conditional transport problem with quadratic cost. In primal form, the problem may be written as
9
or equivalently,
0
The optimizer is 1, and the transport map is the conditional Brenier map
2
which pushes 3 to 4 (Carlier et al., 2014).
The corresponding dual problem is
5
subject to
6
At the optimum, 7 and 8 are convex conjugates in 9 for each 0, and
1
with the gradients acting as mutual inverses almost everywhere (Carlier et al., 2014).
Under injectivity and differentiability, the density relation is governed by conditional Monge–Ampère equations: 2 and symmetrically,
3
These equations tie conditional densities to Jacobians of the forward and inverse maps and formalize the change-of-variables structure induced by the CVQF (Carlier et al., 2014).
This transport-based definition clarifies why scalar quantile regression does not extend directly to multivariate targets. Scalar QR relies on one-dimensional order and pinball loss, whereas CVQR replaces order statistics by monotone transport maps defined as gradients of convex potentials. The monotonicity notion is therefore geometric rather than ordinal (Rosenberg et al., 2022).
3. Linear VQR as the canonical CVQR model
The canonical regression model is the linear CVQF specification, called vector quantile regression. Let 4 collect known transformations of 5, including an intercept. A linear CVQF posits
6
where 7 is a 8 matrix-valued function and, for each 9, the map 0 is the gradient of a convex function in 1: 2 Under correct specification,
3
Convexity of 4 enforces monotonicity of 5 in the multivariate sense (Carlier et al., 2014).
The interpretation of 6 is analogous to scalar quantile regression, but generalized to multivariate ranks. As 7 varies over the reference domain, the columns of 8 trace conditional quantile surfaces of 9 given 0; variation of 1 across 2 reveals heterogeneity of covariate effects across the conditional distribution, while cross-component dependence is carried by the joint map 3 (Carlier et al., 2014).
As 4 becomes richer, the model becomes nonparametric in the sense of series modeling. A sieve specification approximates a smooth convex potential 5 with a tensor-product basis,
6
Under smoothness, 7 and its gradient uniformly approximate 8 and 9 as 0, providing a nonparametric pathway for CVQF estimation. Convexity can be enforced by construction, for example through positive-definite quadratic forms in 1 or convex basis constraints (Carlier et al., 2014).
A frequent misunderstanding is that CVQR estimates individual quantile points independently, as in separate scalar regressions. The framework is explicitly non-local in 2: the monotonicity requirement is a global shape restriction, so the map cannot be estimated pointwise without reference to the full transport structure (Carlier et al., 2014).
4. Identification, misspecification, and the univariate connection
Identification in the foundational framework relies on four classes of conditions: a reference rank law with density and convex support; conditional regularity ensuring existence of inverse ranks; finite second moments for the OT variational characterization; and correctness of the linear CVQF model together with convexity in 3 (Carlier et al., 2014). For feasible implementation and misspecification, the formulation can be relaxed from conditional independence to mean independence. The relaxed primal becomes
4
When the linear model holds and 5 is full rank, this program identifies 6 and 7 (Carlier et al., 2014).
The mean-independence condition is weaker than conditional independence. With rich 8, or saturated 9, they coincide; otherwise the relaxed program targets a quasi-linear representation rather than literal conditional independence of ranks and covariates (Carlier et al., 2014). This distinction becomes central in the analysis beyond correct specification (Carlier et al., 2016).
The paper “Vector quantile regression beyond correct specification” shows that even under misspecification, the VQR problem still has a solution and still yields a global representation of dependence between random vectors (Carlier et al., 2016). In that setting, define
0
Even when 1 is not convex, primal-dual optimality implies
2
where 3 is the convex envelope of 4. The solution therefore replaces a misspecified primitive by its convex envelope and retains a cyclically monotone subgradient representation in 5 (Carlier et al., 2016).
This suggests a natural interpretation of misspecified CVQR as a best convex, monotone approximation generated by the variational system rather than as a failure of the framework. The statement is inferential in phrasing, but it follows the paper’s “best approximation / projection characterization,” in which the convex envelope delivers the closest convex specification induced by the dual constraints (Carlier et al., 2016).
In the univariate case 6, CVQR reduces to the Koenker–Bassett setting with an additional global monotonicity requirement. Under correct specification, it coincides with classical quantile regression. Beyond correct specification, the 2016 paper proves that CVQR is equivalent to Koenker–Bassett quantile regression with a global noncrossing constraint; more precisely, the global monotonicity-constrained Koenker–Bassett program has the same value as the CVQR mean-independence correlation maximization (Carlier et al., 2016). This establishes the scalar theory as a special case rather than an analogy.
5. Estimation and computation
At the population level, the relaxed linear VQR dual problem is an infinite-dimensional linear program: 7 subject to
8
where
9
At the optimum, 00, where 01 defines 02 (Carlier et al., 2014).
In finite samples, one approximates 03 by the empirical distribution over 04 and 05 by a grid 06. The discretized primal linear program is
07
subject to
08
The last constraint enforces mean independence through moment matching. The discretized dual is
09
subject to
10
Efficient implementation leverages sparsity and standard solvers such as Gurobi; the linear program has 11 variables and 12 constraints, where 13 (Carlier et al., 2014).
Subsequent work emphasizes that exact formulations become computationally prohibitive for moderate target dimension, quantile grid size, or feature dimension. “Fast Nonlinear Vector Quantile Regression” extends VQR beyond linear-in-14 parameterizations, introduces vector monotone rearrangement, and proposes fast, GPU-accelerated solvers for linear and nonlinear VQR with fixed memory footprint (Rosenberg et al., 2022). In that paper, the relaxed dual replaces max constraints by a log-sum-exp objective, yielding an unconstrained convex objective for the linear case: 15 As 16, the relaxed dual approaches the exact dual and is equivalent to an entropic-regularized primal (Rosenberg et al., 2022).
The same paper defines a nonlinear specification by introducing a learned feature map 17: 18 This allows lifting or compressing the covariates before fitting the quantile map (Rosenberg et al., 2022). Because relaxed optimization may slightly violate monotonicity, the paper proposes vector monotone rearrangement (VMR), which projects an estimated map onto the set of monotone maps via an OT problem between 19 and the estimated quantile values. In one dimension, this reduces to sorting quantiles to remove crossings (Rosenberg et al., 2022).
The computational trade-off remains explicit throughout the literature. Complexity grows with the product of sample size and rank-grid size, and with target dimension through the 20 quantile grid. GPU acceleration, double mini-batching, and entropic smoothing mitigate but do not remove the curse of dimensionality in 21 (Rosenberg et al., 2022).
6. Applications, extensions, and scope
The foundational empirical application is multiple Engel curve estimation with household expenditure data and a bivariate response consisting of food and housing/heating expenditures. In the one-dimensional case, regressing each component separately via scalar VQR yields quantile curves very close to classical quantile regression, with minimal crossing issues. In the two-dimensional VQR specification,
22
with 23 independent of 24 under correct specification. The fitted surfaces reveal strong own-propensity effects and significant negative cross-covariation at median income, indicating local substitutability between food and housing in that region—an effect not available from separate scalar regressions (Carlier et al., 2014).
The framework has also been extended to non-Euclidean outcome spaces. “Vector Quantile Regression on Manifolds” defines manifold conditional vector quantile functions using Riemannian optimal transport with quadratic geodesic cost and 25-concave potentials. In that setting, the conditional map is
26
and for each 27, it pushes a manifold base law 28 to the conditional law of 29 (Pegoraro et al., 2023). The dual formulation averages the OT losses over 30, uses bounded continuous 31-concave functions in 32, and supports conditional quantile estimation, confidence sets, and likelihood computation on manifolds such as 33 and 34 (Pegoraro et al., 2023).
The manifold paper also derives a conditional likelihood formula through the inverse map,
35
and defines conditional 36-contours by pushing forward base contours centered at a Fréchet mean (Pegoraro et al., 2023). This widens the scope of CVQR from Euclidean multivariate responses to structured geometric domains.
The limitations stated across the literature are consistent. The reference law must be non-atomic with convex support in the Euclidean theory, or appropriate manifold regularity in the geometric theory. Conditional density assumptions are needed for inverse ranks and Monge–Ampère relations, though existence of the forward CVQF does not strictly require continuity of 37. Monotonicity is a global constraint, and relaxed dual solvers may violate it unless corrective procedures such as VMR or involution regularization are used. Alternative reference distributions are possible, but they alter the geometry of the rank space [(Carlier et al., 2014); (Rosenberg et al., 2022); (Pegoraro et al., 2023)].
Taken together, these developments define CVQR as a transport-based theory of conditional multivariate quantiles. Its distinctive commitments are deterministic coupling through 38, monotonicity via convex or 39-concave potentials, and representation of the full conditional distribution rather than selected marginals. In the foundational Euclidean setting, the formal model is VQR; under misspecification, the theory survives through convex-envelope representations; and in recent extensions, the same logic supports nonlinear solvers, monotonicity repair, and regression on manifolds [(Carlier et al., 2014); (Carlier et al., 2016); (Rosenberg et al., 2022); (Pegoraro et al., 2023)].