Randomized Krylov-Schur Eigsolver
- Randomized Krylov-Schur (rKS) is a randomized eigensolver that extracts a few eigenpairs from large matrices using low-dimensional sketching.
- It replaces costly full-dimensional orthogonalization with sketch-based methods, ensuring stable eigenvalue extraction via Schur reordering and restart.
- The method incorporates deflation and locking of converged eigenpairs, accelerating convergence and maintaining backward stability in large-scale eigenproblems.
Searching arXiv for the rKS paper and related randomized orthogonalization references mentioned in the source material. First, locating the main paper by arXiv id. Randomized Krylov–Schur (rKS) is a randomized eigensolver for large-scale eigenvalue problems that seeks a small subset of eigenpairs of a matrix with . It is introduced as a variant of the Krylov–Schur framework in which full-dimensional orthogonalization is replaced by low-dimensional sketching, while Schur truncation, restart, and a practical deflation technique for converged eigenpairs are retained. The method is designed for settings in which only a few largest-magnitude or interior eigenvalues are needed, and a full Schur or QR factorization is impractical. Its defining features are sketch-orthogonalization in a reduced dimension, stable reordering of small Schur factorizations, and locking of converged Ritz vectors so that the eigenspace associated with a prescribed spectral region can be computed efficiently (Damas et al., 7 Aug 2025).
1. Problem formulation and relation to Krylov–Schur
The method addresses the eigenvalue problem
under the assumption that only a small subset of eigenpairs is sought. The motivating examples stated for this regime include structural mechanics, quantum chemistry, and data science, where the dimension is too large for direct dense factorizations and only part of the spectrum is relevant (Damas et al., 7 Aug 2025).
Like classical Krylov methods, rKS operates through low-dimensional projection onto a Krylov subspace,
from which approximate eigenpairs are extracted by Rayleigh–Ritz projection. The description explicitly places rKS in the lineage of the classical Krylov–Schur method of Stewart (2002), which alternates between constructing an Arnoldi-like decomposition of length , computing a small Schur decomposition and truncating unwanted directions, and restarting by re-expanding back to length (Damas et al., 7 Aug 2025).
The principal distinction is computational. In the classical setting, repeated orthonormalization of vectors of length incurs a cost of approximately , while each restart also requires an 0 Schur factorization with cost 1. rKS modifies the orthogonalization stage by maintaining a sketch of the basis in a much lower-dimensional space, so that orthogonality is checked and enforced in the sketch rather than directly in 2. This suggests that rKS should be viewed not as a new spectral projection principle, but as a randomized implementation of the existing Krylov–Schur cycle with a different orthogonalization mechanism (Damas et al., 7 Aug 2025).
2. Randomized sketching and the Arnoldi factorization
The randomization mechanism is formulated through an 3-subspace embedding 4, with 5, such that for every 6 in the relevant 7-dimensional subspace,
8
By Johnson–Lindenstrauss arguments, the description states that a Gaussian or sparse-sign 9 of size 0 has this property with high probability, so orthogonality in the sketch reflects true orthogonality up to 1 distortion (Damas et al., 7 Aug 2025).
The algorithm begins by drawing an oblivious 2-embedding 3 with 4, for example using sparse 5 signs, and by selecting a random unit vector 6 whose sketch is normalized so that 7. It then constructs a length-8 randomized Arnoldi factorization
9
where 0 is sketch-orthonormal and 1 is upper-Hessenberg (Damas et al., 7 Aug 2025).
The parallel sketch variables are
2
At each step 3, the method computes 4, forms its sketch 5, projects in the sketch space via
6
uses the resulting coefficients in the high-dimensional update
7
and normalizes by the sketched norm,
8
After augmentation of 9 and 0, the resulting factorization satisfies
1
with
2
in the stated formulation (Damas et al., 7 Aug 2025).
The reduction in orthogonalization cost is central. Replacing full Gram–Schmidt on 3-vectors by Gram–Schmidt on 4-dimensional sketches changes the orthogonalization cost from 5 to 6, with only an 7 degradation in conditioning. In the paper’s formulation, the randomized Arnoldi basis is therefore controlled through the geometry of the sketch rather than through explicit Euclidean orthogonalization in the ambient space (Damas et al., 7 Aug 2025).
3. Schur contraction, Ritz extraction, and restart
Once the length-8 factorization is available, rKS returns to a standard Krylov–Schur pattern in reduced dimension. The projected Hessenberg matrix 9 is factorized in real Schur form,
0
where 1 is orthonormal and 2 is block-upper triangular. The Schur form is then reordered so that the 3 wanted Ritz values occupy the leading 4 block. Writing 5 with 6, the method contracts the decomposition to
7
so that
8
The small eigenproblem
9
is then solved, and each eigenpair 0 lifts to the Ritz pair
1
The paper’s pseudocode presents the rKS cycle as five stages: expansion to length 2 via randomized Arnoldi; Schur factorization and reordering; Ritz extraction and residual checking; deflation or locking of converged Ritz pairs; and contraction to the unlocked invariant component before the next restart. In that pseudocode, the residual proxy is written as
3
with reference to equation (2.33) in the paper, and Ritz vectors with 4 are locked (Damas et al., 7 Aug 2025).
This design preserves the Arnoldi 5 Schur 6 truncate 7 restart logic of classical Krylov–Schur while moving the most expensive orthogonalization work into the sketch space. A plausible implication is that the method’s behavior remains structurally familiar to practitioners of implicitly restarted Arnoldi and Krylov–Schur, even though the basis management strategy is randomized.
4. Deflation and locking of converged eigenpairs
A distinguishing feature of the method is its practical deflation technique for converged eigenpairs. In the single-vector case, after contraction the decomposition is assumed to have the form
8
with 9. Under this condition, 0 approximately satisfies 1, and the vector is locked by replacing 2 with the deflated operator
3
The description states that 4 is then an exact Schur vector of 5 with eigenvalue 6, while the remaining factorization becomes
7
Subsequent Arnoldi steps enforce orthogonality to the locked vector through
8
The subspace version generalizes this mechanism. If 9 vectors 0 have converged, the projector is formed as
1
and the operator is replaced by
2
The stated consequence is that the locked subspace 3 remains invariant under 4, with zero eigenvalues in 5, while the rKS cycle continues on the complement 6 (Damas et al., 7 Aug 2025).
In operational terms, the deflation mechanism reduces the active subspace dimension after convergence of part of the spectrum. The numerical summary explicitly reports that locking converged vectors reduced subsequent subspace dimension and accelerated convergence, with backward error consistent with theory (Damas et al., 7 Aug 2025). This suggests that deflation is not merely an implementation convenience, but an integral component of the method’s scalability when a sequence of targeted eigenpairs is required.
5. Complexity, conditioning, and stability
The per-restart computational cost is decomposed into matrix–vector products, sketching, low-dimensional orthogonalization, and a small Schur factorization. The paper states the following costs for one cycle: 7 products 8 with complexity 9; sketch formation 0 at 1 per step, or 2 if 3 is very sparse; computation of 4 and related Gram–Schmidt operations in sketch space at 5; and Schur factorization of 6 at 7. The resulting total is summarized as
8
in contrast to the classical cost
9
The stability discussion is likewise phrased in terms of the sketch. Because 00 is an 01-embedding for the evolving Krylov subspace, the sketched basis satisfies
02
Using Corollary 2.2 in Balabanov–Grigori ’22 as cited in the paper, the condition number of the full basis 03 remains 04, which is used to justify that the projected small problem is well-conditioned (Damas et al., 7 Aug 2025).
For a Ritz pair 05 extracted from
06
the residual norm is bounded according to Theorem 2.14 of the paper by
07
The deflation stage introduces an additional tolerance parameter 08, and the paper states that the backward error on each locked eigenpair is 09, so the method is backward stable up to the sketch distortion 10 (Damas et al., 7 Aug 2025).
The stated interpretation is that randomization changes the orthogonalization model but does not forfeit the stability guarantees typically required of restarted Krylov eigensolvers. A plausible implication is that parameter selection for 11, 12, and 13 governs an explicit trade-off between orthogonalization cost and spectral accuracy.
6. Parameterization, implementation, and numerical behavior
The implementation guidance in the paper is specific. The sketch dimension is taken as 14 with 15, and 16 is reported to give 17 in practice. The Krylov dimension is often recommended as 18; if 19 is too small, convergence slows, whereas excessively large 20 increases per-cycle cost. For the embedding 21, the two listed options are dense Gaussian and sparse 22 signs with a few nonzeros per column. For sketch-orthonormalization, the recommended choices are Randomized Gram–Schmidt (RGS) of Balabanov–Grigori ’22 and Randomized Householder QR of Grigori–Timsit ’24, which the description states cost half the flops of classical Gram–Schmidt and maintain sketch-orthonormality to high accuracy. For Schur reordering, LAPACK’s dtrsen or Kressner’s block-reordering algorithms are suggested. Storage requires both 23 and its sketch 24, and contraction applies the same orthonormal transformations to both (Damas et al., 7 Aug 2025).
The numerical experiments summarized in the source description cover both synthetic and real sparse problems. For synthetic tri-diagonal tests with 25 up to 26, 27, and 28, using exponential, logarithmic, harmonic, and geometric diagonal spectra with Gaussian sub-diagonal noise, rKS was reported as 29–30 faster than classical IRA/KS, often converging in fewer restarts, while producing identical Ritz values and residuals below 31. For real sparse matrices from the UF collection, including Atmosmodl and Vas_stokes_4M, with 32 and 33, the method ran 34–35 faster than both MATLAB’s eigs (IRA) and the authors’ deterministic Krylov–Schur implementation, and was described as robust to the choice of 36 (Damas et al., 7 Aug 2025).
These results are presented as evidence that the method scales on large nonsymmetric problems while preserving the overall Krylov–Schur architecture. Within that framing, rKS is characterized by three coupled design choices: low-dimensional sketching in place of full-dimensional orthogonalization, Schur-based restart and truncation on the projected problem, and locking of converged eigenpairs through deflation.