Close the selection lock caused by mixture-based atom scoring
Develop and validate a joint selection criterion for the vocals and accompaniment atom spans that explains the mixture while penalising overlap between the spans, thereby eliminating the cross-source term that causes atoms to be selected for explaining the mixture rather than resembling their respective sources.
References
The lightest remedy, and the closest to the diagnosis, is a joint criterion over both sources: choose $(\mathbf{A}_1, \mathbf{A}_2)$ to explain $\mathbf{x}$ while penalising configurations in which the two spans overlap, that overlap being the mechanism by which explaining the mixture twice beats resembling either source once. A penalty on the principal angles between $\spn \mathbf{A}_1$ and $\spn \mathbf{A}_2$, or equivalently on $|\mathbf{A}_1{*}\mathbf{A}_2|$, expresses that directly and leaves the inference structure untouched. Two approximations of it were measured on the protocol below, one discounting each atom's score by its coherence with the other source's dictionary and one selecting both sets jointly by deflation: at $k = 16$ and 300 atoms they raise the vocals from $+5.57$ to $+6.20$ and $+6.54$~dB against a capacity of $+12.99$, which closes 8 to 13 per cent of the lock and leaves it open.