Skip to main content
Journal cover image

Optimal F-score Matching for Bipartite Record Linkage

Publication ,  Journal Article
Bai, EA; Binette, O; Reiter, JP
Published in: Statistics and Computing
October 1, 2025

Probabilistic record linkage is often used to match records from two files, in particular when the variables common to both files comprise identifiers measured with occasional errors like names and demographic variables. We consider bipartite record linkage settings in which there are no duplicates within the files, but some entities appear in both files. In this setting, the analyst desires a point estimate of the linkage structure that matches each record to at most one record from the other file. We introduce an estimator that maximizes the expected F-score for the linkage structure. We target the approach for record linkage methods that produce either (an approximate) posterior distribution of the unknown linkage structure or probabilities of matches for record pairs. Using simulations and applications with genuine data, we illustrate that the F-score estimators have desirable properties.

Duke Scholars

Published In

Statistics and Computing

DOI

EISSN

1573-1375

ISSN

0960-3174

Publication Date

October 1, 2025

Volume

35

Issue

5

Related Subject Headings

  • Statistics & Probability
  • 4905 Statistics
  • 4903 Numerical and computational mathematics
  • 4901 Applied mathematics
  • 0802 Computation Theory and Mathematics
  • 0104 Statistics
 

Citation

APA
Chicago
ICMJE
MLA
NLM
Bai, E. A., Binette, O., & Reiter, J. P. (2025). Optimal F-score Matching for Bipartite Record Linkage. Statistics and Computing, 35(5). https://doi.org/10.1007/s11222-025-10701-y
Bai, E. A., O. Binette, and J. P. Reiter. “Optimal F-score Matching for Bipartite Record Linkage.” Statistics and Computing 35, no. 5 (October 1, 2025). https://doi.org/10.1007/s11222-025-10701-y.
Bai EA, Binette O, Reiter JP. Optimal F-score Matching for Bipartite Record Linkage. Statistics and Computing. 2025 Oct 1;35(5).
Bai, E. A., et al. “Optimal F-score Matching for Bipartite Record Linkage.” Statistics and Computing, vol. 35, no. 5, Oct. 2025. Scopus, doi:10.1007/s11222-025-10701-y.
Bai EA, Binette O, Reiter JP. Optimal F-score Matching for Bipartite Record Linkage. Statistics and Computing. 2025 Oct 1;35(5).
Journal cover image

Published In

Statistics and Computing

DOI

EISSN

1573-1375

ISSN

0960-3174

Publication Date

October 1, 2025

Volume

35

Issue

5

Related Subject Headings

  • Statistics & Probability
  • 4905 Statistics
  • 4903 Numerical and computational mathematics
  • 4901 Applied mathematics
  • 0802 Computation Theory and Mathematics
  • 0104 Statistics