TítolChoquet integral for record linkage
Publication TypeJournal Article
Year of Publication2012
AuthorsAbril D, Navarro-Arribas G, Torra V
JournalAnnals of Operations Research
EditorSpringer US

Record linkage is used in data privacy to evaluate the disclosure risk of protected data. It models potential attacks, where an intruder attempts to link records from the protected data to the original data. In this paper we introduce a novel distance based record linkage, which uses the Choquet integral to compute the distance between records. We use a fuzzy measure to weight each subset of variables from each record. This allows us to improve standard record linkage and provide insightful information about the re-identification risk of each variable and their interaction. To do that, we use a supervised learning approach which determines the optimal fuzzy measure for the linkage.