Polyjuice: Automated, General-purpose Counterfactual Generation

01/01/2021
by   Tongshuang Wu, et al.
4

Counterfactual examples have been shown to be useful for many applications, including calibrating, evaluating, and explaining model decision boundaries. However, previous methods for generating such counterfactual examples have been tightly tailored to a specific application, used a limited range of linguistic patterns, or are hard to scale. We propose to disentangle counterfactual generation from its use cases, i.e., gather general-purpose counterfactuals first, and then select them for specific applications. We frame the automated counterfactual generation as text generation, and finetune GPT-2 into a generator, Polyjuice, which produces fluent and diverse counterfactuals. Our method also allows control over where perturbations happen and what they do. We show Polyjuice supports multiple use cases: by generating diverse counterfactuals for humans to label, Polyjuice helps produce high-quality datasets for model training and evaluation, requiring 40 When used to generate explanations, Polyjuice helps augment feature attribution methods to reveal models' erroneous behaviors.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
12/20/2022

DISCO: Distilling Phrasal Counterfactuals with Large Language Models

Recent methods demonstrate that data augmentation using counterfactual k...
research
06/21/2022

Plug and Play Counterfactual Text Generation for Model Robustness

Generating counterfactual test-cases is an important backbone for testin...
research
05/26/2023

CREST: A Joint Framework for Rationalization and Counterfactual Text Generation

Selective rationales and counterfactual examples have emerged as two eff...
research
04/24/2023

TIGTEC : Token Importance Guided TExt Counterfactuals

Counterfactual examples explain a prediction by highlighting changes of ...
research
10/10/2022

CORE: A Retrieve-then-Edit Framework for Counterfactual Data Generation

Counterfactual data augmentation (CDA) – i.e., adding minimally perturbe...
research
06/28/2022

Flexible text generation for counterfactual fairness probing

A common approach for testing fairness issues in text-based classifiers ...
research
09/30/2020

Bilateral Asymmetry Guided Counterfactual Generating Network for Mammogram Classification

Mammogram benign or malignant classification with only image-level label...

Please sign up or login with your details

Forgot password? Click here to reset