Scaling DART and evaluating cross-domain effectiveness
Determine whether DART—Depth-as-Target Pretraining for Surgical Vision Foundation Models—remains effective when applied to larger vision-transformer backbones and to domains other than surgical vision.
References
Our evidence is also confined to surgical data and to ViT-{S,B} backbones, so whether DART is effective at larger scales or in other domains remains an open question.
— DART: Depth-as-Target Pretraining for Surgical Vision Foundation Models
(2609.04555 - Han et al., 3 Sep 2026) in Section 5, “Conclusion and Future Work”