Privacy-Preserving Social Media Data Outsourcing Conference Paper uri icon

abstract

  • © 2018 IEEE. User-generated social media data are exploding and of high demand in public and private sectors. The disclosure of intact social media data exacerbates the threats to user privacy. In this paper, we first identify a text-based user-linkage attack on current data outsourcing practices, in which the real users in an anonymized dataset can be pinpointed based on the users' unprotected text data. Then we propose a framework for differentially privacy-preserving social media data outsourcing for the first time in literature. Within our framework, social media data service providers can outsource perturbed datasets to provide users differential privacy while offering high data utility to social media data consumers. Our differential privacy mechanism is based on a novel notion of E - text indistinguishability, which we propose to thwart the text-based user-linkage attack. Extensive experiments on real-world and synthetic datasets confirm that our framework can enable high-level differential privacy protection and also high data utility.

altmetric score

  • 3

author list (cited authors)

  • Zhang, J., Sun, J., Zhang, R., Zhang, Y., & Hu, X.

citation count

  • 15

complete list of authors

  • Zhang, Jinxue||Sun, Jingchao||Zhang, Rui||Zhang, Yanchao||Hu, Xia

publication date

  • April 2018