一种基于聚类的数据匿名方法.PDFVIP

  • 1
  • 0
  • 约5.69万字
  • 约 14页
  • 2019-10-15 发布于天津
  • 举报
ISSN 1000-9825, CODEN RUXUEW E-mail: jos@ Journal of Software , Vol.21, No.4, April 2010, pp.680−693 doi: 10.3724/SP.J.1001.2010.03508 Tel/Fax: +86-10 © by Institute of Software , the Chinese Academy of Sciences . All rights reserved. ∗ 一种基于聚类的数据匿名方法 + 王智慧 , 许 俭, 汪 卫, 施伯乐 (复旦大学 计算机科学技术学院,上海 200433) Clustering-Based Approach for Data Anonymization + WANG Zhi-Hui , XU Jian, WANG Wei, SHI Bai-Le (School of Computer Science, Fudan University, Shanghai 200433, China) + Corresponding author: E-mail: zhhwang@ Wang ZH, Xu J, Wang W, Shi BL. Clustering-Based approach for data anonymization. Journal of Software, 2010,21(4):680−693. /1000-9825/3508.htm Abstract : To prevent the disclosure of privacy, it requires preserving the anonymity of sensitive attributes in data sharing. The attribute values on quasi-identifiers often have to be generalized before data sharing to avoid linking attack, and thus to achieve the anonymity in data sharing. Data generalization increases the uncertainty of attribute values, and results in the loss of information to some extent. Traditional data generalization is often based on the predefined hierarchy, which causes over-generalization and too much unnecessary information loss. In this paper, the attributes in a quasi-identifier are classified into two categories, ordered attributes and unordered attributes. More flexible strategies for data generalization are proposed for them, respectively. At the same time, the loss of information is defined quantitatively based on the change of uncertainty of attribute values during data gen

文档评论(0)

1亿VIP精品文档

相关文档