Disentangle the contributions of temporal modeling, label refinement, and multimodal learning
Determine the individual contributions of temporal modeling, label refinement, and multi-task learning to the macro-phase recognition improvement by conducting dedicated ablation studies.
References
This configuration jointly introduces temporal modeling, label refinement, and multi-task learning relative to the Stage~A baseline; disentangling their individual contributions requires dedicated ablations, which we leave to future work.
— Multimodal Shared Latent Representation of Narration, Microscope and iOCT Images for Phase Recognition in Vitreoretinal Surgery
(2608.31065 - Izmitlioglu et al., 31 Aug 2026) in Section Experimental Results, paragraph beginning “For macro-phase recognition”