When and Why does a Model Fail? A Human-in-the-loop Error Detection Framework for Sentiment Analysis (2106.00954v1)

Published 2 Jun 2021 in cs.CL and cs.HC

Abstract: Although deep neural networks have been widely employed and proven effective in sentiment analysis tasks, it remains challenging for model developers to assess their models for erroneous predictions that might exist prior to deployment. Once deployed, emergent errors can be hard to identify in prediction run-time and impossible to trace back to their sources. To address such gaps, in this paper we propose an error detection framework for sentiment analysis based on explainable features. We perform global-level feature validation with human-in-the-loop assessment, followed by an integration of global and local-level feature contribution analysis. Experimental results show that, given limited human-in-the-loop intervention, our method is able to identify erroneous model predictions on unseen data with high precision.

Citations (9)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

When and Why does a Model Fail? A Human-in-the-loop Error Detection Framework for Sentiment Analysis (2106.00954v1)

Summary

Related Papers