Privacy-preserving Deep Learning based Record Linkage

Ranbaduge, Thilina; Vatsalan, Dinusha; Ding, Ming

Computer Science > Cryptography and Security

arXiv:2211.02161 (cs)

[Submitted on 3 Nov 2022]

Title:Privacy-preserving Deep Learning based Record Linkage

Authors:Thilina Ranbaduge, Dinusha Vatsalan, Ming Ding

View PDF

Abstract:Deep learning-based linkage of records across different databases is becoming increasingly useful in data integration and mining applications to discover new insights from multiple sources of data. However, due to privacy and confidentiality concerns, organisations often are not willing or allowed to share their sensitive data with any external parties, thus making it challenging to build/train deep learning models for record linkage across different organizations' databases. To overcome this limitation, we propose the first deep learning-based multi-party privacy-preserving record linkage (PPRL) protocol that can be used to link sensitive databases held by multiple different organisations. In our approach, each database owner first trains a local deep learning model, which is then uploaded to a secure environment and securely aggregated to create a global model. The global model is then used by a linkage unit to distinguish unlabelled record pairs as matches and non-matches. We utilise differential privacy to achieve provable privacy protection against re-identification attacks. We evaluate the linkage quality and scalability of our approach using several large real-world databases, showing that it can achieve high linkage quality while providing sufficient privacy protection against existing attacks.

Comments:	11 pages
Subjects:	Cryptography and Security (cs.CR); Databases (cs.DB); Data Structures and Algorithms (cs.DS); Information Retrieval (cs.IR); Machine Learning (cs.LG)
Cite as:	arXiv:2211.02161 [cs.CR]
	(or arXiv:2211.02161v1 [cs.CR] for this version)
	https://doi.org/10.48550/arXiv.2211.02161

Submission history

From: Thilina Ranbaduge [view email]
[v1] Thu, 3 Nov 2022 22:10:12 UTC (617 KB)

Computer Science > Cryptography and Security

Title:Privacy-preserving Deep Learning based Record Linkage

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Cryptography and Security

Title:Privacy-preserving Deep Learning based Record Linkage

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators