From outputs to insights: a survey of rationalization approaches for explainable text classification

Research output: Contribution to journalLiterature reviewpeer-review

Abstract

Deep learning models have achieved state-of-the-art performance for text classification in the last two decades. However, this has come at the expense of models becoming less understandable, limiting their application scope in high-stakes domains. The increased interest in explainability has resulted in many proposed forms of explanation. Nevertheless, recent studies have shown that rationales, or language explanations, are more intuitive and human-understandable, especially for non-technical stakeholders. This survey provides an overview of the progress the community has achieved thus far in rationalization approaches for text classification. We first describe and compare techniques for producing extractive and abstractive rationales. Next, we present various rationale-annotated data sets that facilitate the training and evaluation of rationalization models. Then, we detail proxy-based and human-grounded metrics to evaluate machine-generated rationales. Finally, we outline current challenges and encourage directions for future work.
Original languageEnglish
Article number1363531
JournalFrontiers in Artificial Intelligence
Volume7
DOIs
Publication statusPublished - 23 Jul 2024

Keywords

  • rationalisation
  • text classification
  • language explanations

Fingerprint

Dive into the research topics of 'From outputs to insights: a survey of rationalization approaches for explainable text classification'. Together they form a unique fingerprint.

Cite this