Reproducibility and independent explanatory value of definition-conditioned calibration

Determine whether the observed improvement from adding learner-definition terms and definition–score interactions to sensor-derived engagement calibration reproduces in larger cohorts and whether learner definitions explain prediction gains independently of other participant or cohort differences.

Background

The definition-conditioned calibration model used only eighteen learners and included twenty design columns, comprising sensor-derived scores, definition indicators, and definition–score interactions. The paper reports an exploratory performance gain, but the ratio of model complexity to the number of independent learners limits confidence in its stability. The authors explicitly note that the analysis does not establish either reproducibility of the gain or an independent explanatory role for the definitions.

References

Holding upstream scores fixed makes the comparison specific to the added definition terms, but does not establish that the gain will reproduce or that definitions explain it independently of other participant or cohort differences.

— E3Sense: Head-Confined Multimodal Sensing of Learner Engagement  (2609.26569 - Anupkrishnan et al., 22 Sep 2026) in Section 7.4, Discussion—Definition Elicitation and Model Complexity