A4NT: Author Attribute Anonymity by Adversarial Training of Neural Machine Translation
conference contribution
posted on 2023-11-29, 18:08authored byRakshith Shetty, Bernt Schiele, Mario FritzMario Fritz
Text-based analysis methods enable an adversary to reveal privacy relevant author attributes such as gender, age and can identify the text's author. Such methods can compromise the privacy of an anonymous author even when the author tries to remove privacy sensitive content. In this paper, we propose an automatic method, called the Adversarial Author Attribute Anonymity Neural Translation ($\text{A}^{4}\text{NT}$), to combat such text-based adversaries. Unlike prior works on obfuscation, we propose a system that is fully automatic and learns to perform obfuscation entirely from the data. This allows us to easily apply the $\text{A}^{4}\text{NT}$ system to obfuscate different author attributes. We propose a sequence-to-sequence language model, inspired by machine translation, and an adversarial training framework to design a system which learns to transform the input text to obfuscate the author attributes without paired data. We also propose and evaluate techniques to impose constraints on our $\text{A}^{4}\text{NT}$ model to preserve the semantics of the input text. $\text{A}^{4}\text{NT}$ learns to make minimal changes to the input to successfully fool author attribute classifiers, while preserving the meaning of the input text. Our experiments on two datasets and three settings show that the proposed method is effective in fooling the attribute classifiers and thus improves the anonymity of authors.
History
Preferred Citation
Rakshith Shetty, Bernt Schiele and Mario Fritz. A4NT: Author Attribute Anonymity by Adversarial Training of Neural Machine Translation. In: Usenix Security Symposium (USENIX-Security). 2018.
Primary Research Area
Trustworthy Information Processing
Name of Conference
Usenix Security Symposium (USENIX-Security)
Legacy Posted Date
2018-07-06
Open Access Type
Gold
BibTeX
@inproceedings{cispa_all_2628,
title = "A4NT: Author Attribute Anonymity by Adversarial Training of Neural Machine Translation",
author = "Shetty, Rakshith and Schiele, Bernt and Fritz, Mario",
booktitle="{Usenix Security Symposium (USENIX-Security)}",
year="2018",
}