Evaluate CGMIA on larger language models
Determine whether CGMIA remains effective for larger code-generation language models with substantially more parameters than the smaller models used in the study.
References
Consequently, the effectiveness of CGMIA on larger models with substantially more parameters remains uncertain and requires further investigation.
— Keep Evaluation Fair: Detecting Data Leakage in Code Generation Benchmarks via Membership Inference Attacks
(2609.09865 - Zhao et al., 9 Sep 2026) in Section 10, Threats of validity, item (1)