Understanding and Mitigating Spurious Correlations in Text Classification with Neighborhood Analysis (2305.13654v3)

Published 23 May 2023 in cs.CL

Abstract: Recent research has revealed that machine learning models have a tendency to leverage spurious correlations that exist in the training set but may not hold true in general circumstances. For instance, a sentiment classifier may erroneously learn that the token "performances" is commonly associated with positive movie reviews. Relying on these spurious correlations degrades the classifiers performance when it deploys on out-of-distribution data. In this paper, we examine the implications of spurious correlations through a novel perspective called neighborhood analysis. The analysis uncovers how spurious correlations lead unrelated words to erroneously cluster together in the embedding space. Driven by the analysis, we design a metric to detect spurious tokens and also propose a family of regularization methods, NFL (doN't Forget your Language) to mitigate spurious correlations in text classification. Experiments show that NFL can effectively prevent erroneous clusters and significantly improve the robustness of classifiers without auxiliary data. The code is publicly available at https://github.com/oscarchew/doNt-Forget-your-Language.

Authors (4)

Oscar Chew (2 papers)
Hsuan-Tien Lin (43 papers)
Kai-Wei Chang (292 papers)
Kuan-Hao Huang (33 papers)

Citations (4)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

GitHub

GitHub - oscarchew/doNt-Forget-your-Language: The official repository of our EACL2024-Findings paper: Understanding and Mitigating Spurious Correlations in Text Classification with Neighborhood Analysis

Tweets

https://twitter.com/kuanhaoh_/status/1754995504558071920

https://twitter.com/hsuantienlin/status/1748814592351015192

Understanding and Mitigating Spurious Correlations in Text Classification with Neighborhood Analysis (2305.13654v3)

Summary

Related Papers

GitHub

Tweets