Skip to main navigation Skip to search Skip to main content

Better Hit the Nail on the Head than Beat around the Bush: Removing Protected Attributes with a Single Projection

Research output: Chapter in Book / Report / Conference proceedingConference contributionAcademicpeer-review

Abstract

Bias elimination and recent probing studies attempt to remove specific information from embedding spaces. Here it is important to remove as much of the target information as possible, while preserving any other information present. INLP is a popular recent method which removes specific information through iterative nullspace projections. Multiple iterations, however, increase the risk that information other than the target is negatively affected. We introduce two methods that find a single targeted projection: Mean Projection (MP, more efficient) and Tukey Median Projection (TMP, with theoretical guarantees). Our comparison between MP and INLP shows that (1) one MP projection removes linear separability based on the target and (2) MP has less impact on the overall space. Further analysis shows that applying random projections after MP leads to the same overall effects on the embedding space as the multiple projections of INLP. Applying one targeted (MP) projection hence is methodologically cleaner than applying multiple (INLP) projections that introduce random effects.
Original languageEnglish
Title of host publicationProceedings of the 2022 Conference on Empirical Methods in Natural Language Processing
PublisherACL Anthology
Pages8395-8416
Number of pages22
Publication statusPublished - 2022

Fingerprint

Dive into the research topics of 'Better Hit the Nail on the Head than Beat around the Bush: Removing Protected Attributes with a Single Projection'. Together they form a unique fingerprint.

Cite this