Source of incomplete masks in prompted SAM pipelines

Determine whether the incomplete bruise masks produced by bounding-box localisation followed by SAM-style segmentation arise from bounding-box localisation, prompting, or mask generation.

Background

The evaluation compares several large vision-language-model prompting pipelines paired with SAM2 or MedSAM. These systems generally produce conservative or incomplete bruise masks, particularly because diffuse bruises are difficult to localise and delineate using bounding-box and prompt-based workflows.

Although the results establish that these pipelines underperform the proposed BruNet variants, the evaluation does not disentangle the separate contributions of the initial bounding-box localisation, the prompting process, and the subsequent mask-generation stage. Identifying the source of the limitation remains unresolved.

References

However, the current evaluation does not isolate whether this limitation originates from bounding-box localisation, prompting, or mask generation.

BruNet: A Cross-Domain Transfer Framework for Bruise Segmentation  (2609.11463 - Wang et al., 10 Sep 2026) in Section 4.1, “Results”