Identify the mechanism underlying stable gaze–AOI divergence under combined guidance
Determine whether stable gaze–AOI divergence during combined Co-Annotator guidance occurs because the ontology-bounded VLM text satisfies residents’ information needs or because residents rely less on the gaze-aligned AOI overlay when both modalities are presented.
References
Two interpretations are consistent with this pattern: the VLM text may have satisfied residents' information needs, reducing pressure to spatially re-orient gaze; or residents may have relied less on the AOI overlay in the combined condition, engaging primarily with the textual draft. Survey data showing moderate perceived reliance (3.75/7) and qualitative reports of a corroboration loop between modalities favor the first interpretation, though direct measurement of per-modality engagement is needed to distinguish the two.