Tensor Nuclear Norm Overview
- Tensor nuclear norm is a generalization of the matrix nuclear norm defined via unit rank-one tensor decompositions and serves as the dual of the tensor spectral norm.
- The t-product based tensor nuclear norm offers an exact convex envelope of tensor average rank, underpinning recovery theory in robust PCA and tensor completion.
- Variants such as p-TNN, truncated TNN, and framelet-based TNN refine low-rank approximations, targeting specific data geometries and improving computational efficiency.
Tensor nuclear norm denotes several extensions of the matrix nuclear norm to multiway arrays. In the generic higher-order setting, it is the dual of the tensor spectral norm and can be written as the minimum total weight in a decomposition into unit rank-one tensors (Friedland et al., 2014). In the third-order -SVD literature, the term often refers to the norm induced by the tensor–tensor product, defined through a block-circulant embedding or, equivalently, by Fourier-domain frontal slices; within that framework it is the convex envelope of tensor average rank on the unit ball of the tensor spectral norm (Lu et al., 2018). The literature also contains several other tensor nuclear-norm constructions and surrogates, including truncated, transform-based, nonlinear, tensor-ring, and semidefinite-relaxation variants, reflecting the fact that low-rank structure in tensors is model-dependent (Xue et al., 2019).
1. General formulations and relation to matrices
For an order- tensor , one standard definition is
with dual spectral norm
When , this reduces exactly to the matrix nuclear norm (Friedland et al., 2014). Qi et al. state the same decomposition-based definition for general tensors and emphasize the duality (Qi et al., 2019).
This formulation preserves several matrix-like features, but not all. Every tensor has a nuclear norm attaining decomposition, and every symmetric tensor has a symmetric nuclear norm attaining decomposition; for symmetric tensors, the symmetric nuclear norm equals the nuclear norm (Friedland et al., 2014). Nie further develops moment-SOS and Lasserre-relaxation machinery for computing symmetric tensor nuclear norms and extracting symmetric nuclear decompositions in moderate-scale cases (Nie, 2016). At the same time, higher-order behavior departs sharply from matrices: for , there exist real tensors whose real and complex nuclear norms differ, and exact computation is NP-hard in several senses (Friedland et al., 2014).
A recurring source of ambiguity is that the literature contains several definitions of tensor nuclear norm for third-order tensors, especially in completion and recovery. The -SVD-based definition, the sum of unfolding nuclear norms, tensor-ring nuclear norms, and semidefinite 0-norm relaxations all target low-rankness, but they act on different tensor models and induce different optimization geometries (Xue et al., 2017).
2. The 1-product-based tensor nuclear norm
The construction introduced in tensor robust PCA is built on the tensor–tensor product. For 2 and 3,
4
where 5 is the 6 block-circulant matrix built from the frontal slices of 7. Equivalently, one may FFT each tensor along the third dimension, multiply the resulting block-diagonal matrices slice-by-slice, and then inverse-FFT (Lu et al., 2018).
This algebra supports a 8-SVD
9
with 0 orthogonal under the 1-product and 2 3-diagonal. If 4, the tensor spectral norm and tensor nuclear norm are defined by
5
In the Fourier-domain viewpoint, if 6 has frontal slices 7, then
8
These definitions extend the matrix case exactly when 9 (Lu et al., 2018).
The same framework defines the tensor average rank
0
In related 1-SVD formulations, one also encounters tubal rank, defined as the number of nonzero singular tubes or, equivalently, 2 after FFT along the third mode (Zhang et al., 2022). The distinction matters: average-rank statements and tubal-rank statements are not interchangeable unless a paper states the relevant equivalence or reduction.
3. Convex-envelope geometry and recovery theory
A central theorem of the 3-product-based construction is that on the unit ball
4
the convex envelope of 5 is exactly 6 (Lu et al., 2018). The proof proceeds through convex conjugates. Writing 7, the conjugate 8 can be expressed through the singular values of the block-circulant embedding, and von Neumann’s trace inequality yields the optimal alignment argument. The biconjugate 9, restricted to 0, then reduces to
1
This reproduces the matrix relation between nuclear norm, spectral norm, and rank in a tensor algebra induced by the 2-product (Lu et al., 2018).
That convex-envelope statement is specific to the 3-product model. The same paper stresses that, unlike just summing matrix nuclear norms of all slices, the new tensor nuclear norm averages over the block-circulant embedding, ensuring the convex-envelope property holds. In this sense, the 4-SVD-based TNN is not merely a heuristic slice regularizer but a precise convex relaxation of tensor average rank in the associated non-commutative algebra (Lu et al., 2018).
Recovery theory in the same framework extends beyond robust PCA. By choosing an atomic set adapted to tubal rank, Zhang and Aeron show that TNN is a special atomic norm and that exact recovery from Gaussian measurements of a tensor of size 5 and tubal rank 6 requires
7
which is order optimal relative to the degrees of freedom 8 (Lu et al., 2018). The same work gives tensor-completion guarantees under uniform random sampling and incoherence assumptions, with sample complexity 9 (Lu et al., 2018).
In tensor robust PCA, the 0-product-based TNN leads to a convex program that exactly recovers low-rank and sparse components under incoherence and sparsity assumptions, paralleling matrix PCP theory. Matrix RPCA appears as a special case, and the reported applications include image recovery and background modeling (Lu et al., 2018).
4. Computation, proximal mappings, and algorithmic use
In the 1-SVD setting, computing the TNN and its proximal operator is FFT-centric. One performs an FFT along mode 3, computes matrix SVDs on the frontal slices in the transform domain, applies slice-wise singular-value thresholding, and then returns by inverse FFT. For the proximal map of 2, the tensor singular value thresholding (TSVT) procedure consists of: FFT, 3 slice SVDs, soft-thresholding of singular values, and inverse FFT (Zhang et al., 2022).
The per-iteration complexity reported for TSVT is one FFT and one inverse FFT of size 4, together with 5 independent SVDs of size 6 (Zhang et al., 2022). This decomposition is the basic computational reason why ADMM and proximal-gradient methods are practical for large third-order tensors in imaging and completion.
A direct application is dynamic cardiac MRI reconstruction. The TMNN model combines the 7-SVD-based TNN with the Casorati matrix nuclear norm: 8 The motivation is explicit: TNN exploits spatial structure of the dynamic MR data, while the Casorati matrix nuclear norm exploits temporal correlation. The resulting ADMM solver admits a fast Cartesian-sampling variant, and the reported experiments show up to 9 dB SNR gain over plain MNN together with an 0 run-time reduction for the fast 1-space implementation (Zhang et al., 2022).
The same study also states a useful structural caveat: in general,
2
Accordingly, TNN and Casorati nuclear norms should be regarded as complementary regularizers rather than interchangeable ones (Zhang et al., 2022).
5. Variants, tighter surrogates, and transform-induced extensions
Several later works keep the 3-SVD backbone but modify the singular-value penalty or the transform domain to approximate tensor rank more tightly.
Tensor 4-shrinkage nuclear norm. The 5-TNN replaces soft-thresholding by the scalar 6-shrinkage operator
7
applied to the diagonal entries of the 8-SVD core in the Fourier domain. The resulting functional is positive, unitary-invariant, and non-convex for 9, and it satisfies
0
The paper states that 1-TNN is a better approximation of tensor average rank than TNN when 2, gives a statistical recovery-error bound for low-rank tensor completion, and analyzes an ADMM solver with adaptive momentum whose convergence rate is 3 under the smoothness assumption (Liu et al., 2019).
Tensor truncated nuclear norm. T-TNN generalizes matrix truncated nuclear norm to the 4-SVD setting by subtracting the contribution of the leading 5 singular values. In one formulation,
6
so only the tail singular values of the first Fourier frontal slice are penalized. The stated motivation is that truncation avoids over-shrinkage of dominant components and more closely approximates tubal rank than the plain T-NN. The ADMM and APGL schemes require one matrix SVD of size 7 per iteration, and the reported image-completion experiments show a 8–9 speedup over several baselines (Xue et al., 2017).
Framelet-based tensor nuclear norm. F-TNN replaces the DFT along each tube by a redundant framelet transform 0 satisfying 1. The norm is then defined as the sum of matrix nuclear norms of the framelet-transformed frontal slices. Because of framelet-basis redundancy, the representation of each tube is sparsely represented, and the paper reports that on MRI, video, and multispectral data the average truncated frontal-slice rank can drop by 2–3. The completion model is convex, global minimizers can be obtained, and empirical gains of 4–5 dB in PSNR over several baselines are reported (Jiang et al., 2019).
Nonlinear transform induced tensor nuclear norm. NTTNN augments a semi-orthogonal linear transform along mode 3 with an element-wise nonlinear activation 6, giving
7
The paper argues that the nonlinearity can shrink small entries and emphasize dominant modes in the transformed frontal slices, and it solves the resulting nonlinear, nonconvex completion model with proximal alternating minimization under a Kurdyka–Łojasiewicz analysis. Reported gains are 8–9 dB PSNR over TNN and 0–1 dB over learned-transform TNNs on hyperspectral images, multispectral images, and videos (Li et al., 2021).
Multimode nonlinear transform-based TNN. MNT-TNN extends single-mode TTNN to a multimode setting using a face-wise mode-2 transform 3, a mode-4 transform 5, a mode-3 transform 6, and an element-wise nonlinearity 7. The paper formulates a PAM algorithm with sufficient-decrease and relative-error guarantees under the KŁ framework, and introduces ATTNN chains that first use linear TTNN variants and then nonlinear TTNN variants. At very high missing rates, ATTNNs are reported to deliver 8–9 further MAPE/RMSE improvements over standalone MNT-TNN in spatiotemporal traffic imputation (Lu et al., 29 Mar 2025).
Tensor-ring nuclear norm. A different line of work defines
00
where 01 are circular unfoldings. The stated theoretical link is 02 when 03 has TR ranks 04. The resulting completion model is convex and is solved by ADMM; in stripe-missing image and video completion it is reported to outperform several conventional tensor-completion methods (Yu et al., 2019).
6. Higher-order structure, bounds, relaxations, and broader roles
For generic higher-order tensors, structural and computational questions remain difficult. Friedland and Lim show that nuclear norm depends on the base field for tensors of order at least three, that the nuclear norm unit ball and its weak-membership problem are computationally hard, and that computing spectral or nuclear norm is NP-hard for several restricted tensor classes (Friedland et al., 2014). Qi et al. add that the 05-norm, Frobenius norm, and nuclear norm are tensor norms in the sense of vector-norm axioms plus submultiplicativity under outer products, whereas the infinity norm and spectral norm are not tensor norms (Qi et al., 2019).
Despite the intractability of exact computation, there are computable bounds. For an 06-tensor 07, the nuclear norm of every matrix flattening 08 is a lower bound for 09, and for 3-tensors one has
10
with both bounds sharp when 11 (Hu, 2014). For third-order tensors, contraction to positive semidefinite biquadratic tensors gives additional lower bounds: the square roots of the nuclear norms of the three contracted biquadratic tensors are lower bounds of the tensor nuclear norm (Qi et al., 2019).
Semidefinite relaxations provide another route. The 12-norms are defined through theta bodies of polynomial ideals generated by second-order minors of tensor matricizations, with the property that in the matrix case they reduce to the nuclear norm, while for order 13 they give new norms. The unit-14-norm balls converge asymptotically to the unit tensor nuclear norm ball, and computing 15-norms or minimizing them under affine constraints reduces to semidefinite programming (Rauhut et al., 2015).
Recent work on decomposability and subdifferentials shows that the tensor nuclear norm admits full decomposability over specific subspaces such as
16
and identifies the largest possible subspaces allowing exact additivity. The same work derives new inclusions for the subdifferential and uses them to establish the statistical performance of tensor robust PCA for tensors of arbitrary order, described there as the first such result in that generality (Guan et al., 6 Oct 2025).
The tensor nuclear norm also appears outside completion and denoising. In bilinear complexity, a bilinear operator 17 can be identified with a 3-tensor, and its nuclear norm
18
equals the minimum growth factor of any bilinear algorithm for 19. The forward-error bound proved in that setting ties numerical accuracy directly to 20, yielding a tensorial notion of bilinear stability that is invariant under orthogonal change of coordinates (Dai et al., 2022).
Taken together, these results show that “tensor nuclear norm” is not a single universally adopted object but a family of norms and norm-like surrogates attached to distinct tensor models. The decomposition-based higher-order norm provides the broad convex analogue of matrix nuclear norm; the 21-SVD-based norm provides an exact convex envelope of tensor average rank in the 22-product algebra; and later transform-based or truncated variants modify the penalty to target specific data geometries or rank surrogates more tightly (Friedland et al., 2014).