Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation

Bhattacharjee, Amrita; Moraffah, Raha; Garland, Joshua; Liu, Huan

Computer Science > Computation and Language

arXiv:2405.04793 (cs)

[Submitted on 8 May 2024 (v1), last revised 19 Nov 2024 (this version, v2)]

Title:Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation

Authors:Amrita Bhattacharjee, Raha Moraffah, Joshua Garland, Huan Liu

View PDF HTML (experimental)

Abstract:With the development and proliferation of large, complex, black-box models for solving many natural language processing (NLP) tasks, there is also an increasing necessity of methods to stress-test these models and provide some degree of interpretability or explainability. While counterfactual examples are useful in this regard, automated generation of counterfactuals is a data and resource intensive process. such methods depend on models such as pre-trained language models that are then fine-tuned on auxiliary, often task-specific datasets, that may be infeasible to build in practice, especially for new tasks and data domains. Therefore, in this work we explore the possibility of leveraging large language models (LLMs) for zero-shot counterfactual generation in order to stress-test NLP models. We propose a structured pipeline to facilitate this generation, and we hypothesize that the instruction-following and textual understanding capabilities of recent LLMs can be effectively leveraged for generating high quality counterfactuals in a zero-shot manner, without requiring any training or fine-tuning. Through comprehensive experiments on a variety of propreitary and open-source LLMs, along with various downstream tasks in NLP, we explore the efficacy of LLMs as zero-shot counterfactual generators in evaluating and explaining black-box NLP models.

Comments:	Longer version of short paper accepted at IEEE BigData 2024 (Main Track)
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2405.04793 [cs.CL]
	(or arXiv:2405.04793v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2405.04793

Submission history

From: Amrita Bhattacharjee [view email]
[v1] Wed, 8 May 2024 03:57:45 UTC (575 KB)
[v2] Tue, 19 Nov 2024 10:59:30 UTC (1,106 KB)

Computer Science > Computation and Language

Title:Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators