Mutual-information estimation without the sufficient encoder assumption
Estimate or approximate the mutual information I(X;\tilde{X}) for SMILE's selected multimodal medical representations without relying on the sufficient encoder assumption, thereby addressing the information-theoretic difficulty of performing this estimation directly.
References
First, the sufficient encoder assumption in (\ref{eq: multi_loss2}) simplifies optimization but remains over-optimistic; estimating or approximating $I(X;\tilde{X})$ without it is an open problem, even from an information theory perspective.
— SMILE: Self-Explainable Multimodal Information Bottleneck for Medical Diagnosis
(2609.05174 - Yang et al., 4 Sep 2026) in Section Conclusion and Future Work