Use of graph theory measures to identify errors in record linkage

Randall, Sean; Boyd, James; Ferrante, Anna; Bauer, J.; Semmens, James

doi:10.1016/j.cmpb.2014.03.008

dc.contributor.author	Randall, Sean
dc.contributor.author	Boyd, James
dc.contributor.author	Ferrante, Anna
dc.contributor.author	Bauer, J.
dc.contributor.author	Semmens, James
dc.date.accessioned	2017-01-30T10:29:23Z
dc.date.available	2017-01-30T10:29:23Z
dc.date.created	2014-07-03T20:00:23Z
dc.date.issued	2014
dc.identifier.citation	Randall, S. and Boyd, J. and Ferrante, A. and Bauer, J. and Semmens, J. 2014. Use of graph theory measures to identify errors in record linkage. Computer Methods and Programs in Biomedicine. 115 (2): pp. 55-63.
dc.identifier.uri	http://hdl.handle.net/20.500.11937/3205
dc.identifier.doi	10.1016/j.cmpb.2014.03.008
dc.description.abstract	Ensuring high linkage quality is important in many record linkage applications. Current methods for ensuring quality are manual and resource intensive. This paper seeks to determine the effectiveness of graph theory techniques in identifying record linkage errors. A range of graph theory techniques was applied to two linked datasets, with known truth sets. The ability of graph theory techniques to identify groups containing errors was compared to a widely used threshold setting technique. This methodology shows promise; however, further investigations into graph theory techniques are required. The development of more efficient and effective methods of improving linkage quality will result in higher quality datasets that can be delivered to researchers in shorter timeframes.
dc.publisher	Elsevier Ireland Ltd
dc.subject	Record linkage
dc.subject	Graph theory
dc.subject	Data quality
dc.title	Use of graph theory measures to identify errors in record linkage
dc.type	Journal Article
dcterms.source.volume	115
dcterms.source.number	2
dcterms.source.startPage	55
dcterms.source.endPage	63
dcterms.source.issn	01692607
dcterms.source.title	Computer Methods and Programs in Biomedicine
curtin.note	NOTICE: This is the author’s version of a work that was accepted for publication in Computer Methods and Programs in Biomedicine. Changes resulting from the publishing process, such as peer review, editing, corrections, structural formatting, and other quality control mechanisms may not be reflected in this document. Changes may have been made to this work since it was submitted for publication. A definitive version was subsequently published in Computer Methods and Programs in Biomedicine, Vol. 115, Issue 2. (2014). doi: 10.1016/j.cmpb.2014.03.008
curtin.department
curtin.accessStatus	Open access

Files in this item

Name:: 199679_199679.pdf
Size:: 585.2Kb
Format:: PDF

This item appears in the following Collection(s)

Curtin Research Publications

Show simple item record

Use of graph theory measures to identify errors in record linkage

Files in this item

This item appears in the following Collection(s)

Related items