Membership Inference Attacks on Sequence Models (2506.05126v1)

Published 5 Jun 2025 in cs.CR and cs.LG

Abstract: Sequence models, such as LLMs and autoregressive image generators, have a tendency to memorize and inadvertently leak sensitive information. While this tendency has critical legal implications, existing tools are insufficient to audit the resulting risks. We hypothesize that those tools' shortcomings are due to mismatched assumptions. Thus, we argue that effectively measuring privacy leakage in sequence models requires leveraging the correlations inherent in sequential generation. To illustrate this, we adapt a state-of-the-art membership inference attack to explicitly model within-sequence correlations, thereby demonstrating how a strong existing attack can be naturally extended to suit the structure of sequence models. Through a case study, we show that our adaptations consistently improve the effectiveness of memorization audits without introducing additional computational costs. Our work hence serves as an important stepping stone toward reliable memorization audits for large sequence models.

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Tweets

https://twitter.com/chaumian/status/1931460411876077874

https://twitter.com/FSFG/status/1931152050030571528

Membership Inference Attacks on Sequence Models (2506.05126v1)

Summary

Related Papers

Tweets