Performance portability across GPU vendors for irregular kernels
Establish performance portability across GPU vendors for irregular kernels, achieving parity performance from a single source across NVIDIA CUDA and AMD ROCm/HIP targets.
References
Performance portability across vendors---one source, parity performance on each target---remains substantially unsolved.
— Bridging the Vendor Gap: Enabling AMD GPU Support for Awkward Array via ROCm/HIP for the HL-LHC Era
(2609.24628 - Osborne et al., 21 Sep 2026) in Section 2, “Background and motivation,” subsection “Why GPUs, and why portability”