Performance of gender detection tools: a comparative study of name-to-gender inference services-Reference-Cited by-同舟云学术

Performance of gender detection tools: a comparative study of name-to-gender inference services

Published:2021-10-05 Issue:3 Volume:109 Page:
ISSN:1558-9439
Container-title:Journal of the Medical Library Association
language:
Short-container-title:jmla

Author:

Sebo Paul

Abstract

Objective: To evaluate the performance of gender detection tools that allow the uploading of files (e.g., Excel or CSV files) containing first names, are usable by researchers without advanced computer skills, and are at least partially free of charge.Methods: The study was conducted using four physician datasets (total number of physicians: 6,131; 50.3% female) from Switzerland, a multilingual country. Four gender detection tools met the inclusion criteria: three partially free (Gender API, NamSor, and genderize.io) and one completely free (Wiki-Gendersort). For each tool, we recorded the number of correct classifications (i.e., correct gender assigned to a name), misclassifications (i.e., wrong gender assigned to a name), and nonclassifications (i.e., no gender assigned). We computed three metrics: the proportion of misclassifications excluding nonclassifications (errorCodedWithoutNA), the proportion of nonclassifications (naCoded), and the proportion of misclassifications and nonclassifications (errorCoded).Results: The proportion of misclassifications was low for all four gender detection tools (errorCodedWithoutNA between 1.5 and 2.2%). By contrast, the proportion of unrecognized names (naCoded) varied: 0% for NamSor, 0.3% for Gender API, 4.5% for Wiki-Gendersort, and 16.4% for genderize.io. Using errorCoded, which penalizes both types of error equally, we obtained the following results: Gender API 1.8%, NamSor 2.0%, Wiki-Gendersort 6.6%, and genderize.io 17.7%.Conclusions: Gender API and NamSor were the most accurate tools. Genderize.io led to a high number of nonclassifications. Wiki-Gendersort may be a good compromise for researchers wishing to use a completely free tool. Other studies would be useful to evaluate the performance of these tools in other populations (e.g., Asian).

Publisher

University Library System, University of Pittsburgh

Subject

Library and Information Sciences,Health Informatics

Cited by 97 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Investigating the causal effects of affiliation diversity on the disruption of papers in Artificial Intelligence;Information Processing & Management;2024-09

2. Trends and Influences in women authorship of randomized controlled trials in rheumatology: a comprehensive analysis of all published RCTs from 2009 to 2023;2024-08-26

3. Quantifying attrition in science: a cohort-based, longitudinal study of scientists in 38 OECD countries;Higher Education;2024-08-23

4. Twenty-Five Years of Progress—Lessons Learned From JMIR Publications to Address Gender Parity in Digital Health Authorships: Bibliometric Analysis;Journal of Medical Internet Research;2024-08-09

5. Emerging leaders or persistent gaps? Generative AI research may foster women in STEM;International Journal of Information Management;2024-08