OALib Journal期刊
ISSN: 2333-9721
费用：99美元

投递稿件

查看量	下载量

相关文章
更多...

软件学报 2010

Clustering-Based Approach for Data Anonymization
一种基于聚类的数据匿名方法

WANG Zhi-Hui,XU Jian,WANG Wei,SHI Bai-Le,
王智慧,许俭,汪卫,施伯乐

Keywords: data anonymization,quasi-identifier,linking attack,clustering,information loss
数据匿名,准标识符,链接攻击,聚类,信息损失

Full-Text Cite this paper Add to My Lib

Abstract:

To prevent the disclosure of privacy, it requires preserving the anonymity of sensitive attributes in data sharing. The attribute values on quasi-identifiers often have to be generalized before data sharing to avoid linking attack, and thus to achieve the anonymity in data sharing. Data generalization increases the uncertainty of attribute values, and results in the loss of information to some extent. Traditional data generalization is often based on the predefined hierarchy, which causes over-generalization and too much unnecessary information loss. In this paper, the attributes in a quasi-identifier are classified into two categories, ordered attributes and unordered attributes. More flexible strategies for data generalization are proposed for them, respectively. At the same time, the loss of information is defined quantitatively based on the change of uncertainty of attribute values during data generalization. Furthermore, data anonymization is modeled by a clustering problem with special constraints. A clustering-based approach, called L-clustering, is presented for the l-diversity model. L-clustering can meet the requirement of preserving anonymity of sensitive attributes in data sharing, and reduce greatly the amount of information loss resulting from data generalization for implementing data anonymization.

Full-Text

Contact Us

service@oalib.com

QQ:3279437679

WhatsApp +8615387084133

Clustering-Based Approach for Data Anonymization一种基于聚类的数据匿名方法

Clustering-Based Approach for Data Anonymization
一种基于聚类的数据匿名方法