Debiasing Without Protected Attributes: Latent Concept Erasure from Textual Profiles
无保护属性的去偏:从文本画像中消除潜在概念
机构 * University of Cambridge(剑桥大学) ; University of Edinburgh(爱丁堡大学) ; University of Groningen(格罗宁根大学) ; NVIDIA Research(英伟达研究院)
AI总结 提出H-SAL方法,利用自我描述文本作为隐式信号进行后处理概念和属性消除,在无直接敏感属性下实现去偏,并在多领域Stack Exchange基准上验证其效果与显式标签去偏相当或更优。
Comments 23 pages, 5 figures, 12 tables. The paper is currently under review