Sensitivity of DART representations to pseudo-depth quality
Characterize how sensitive the representations learned by DART—Depth-as-Target Pretraining for Surgical Vision Foundation Models—to the quality of pseudo-labeled depth maps and to the choice of the monocular depth estimator used to generate them.
References
We do not characterize how sensitive the learned representations are to depth quality or the choice of the estimator.
— DART: Depth-as-Target Pretraining for Surgical Vision Foundation Models
(2609.04555 - Han et al., 3 Sep 2026) in Section 5, “Conclusion and Future Work”