Gender bias in machine learning for sentiment analysis-Reference-Cited by-同舟云学术

Gender bias in machine learning for sentiment analysis

Published:2018-06-11 Issue:3 Volume:42 Page:343-354
ISSN:1468-4527
Container-title:Online Information Review
language:en
Short-container-title:OIR

Author:

Thelwall Mike^ORCID

Abstract

Purpose The purpose of this paper is to investigate whether machine learning induces gender biases in the sense of results that are more accurate for male authors or for female authors. It also investigates whether training separate male and female variants could improve the accuracy of machine learning for sentiment analysis. Design/methodology/approach This paper uses ratings-balanced sets of reviews of restaurants and hotels (3 sets) to train algorithms with and without gender selection. Findings Accuracy is higher on female-authored reviews than on male-authored reviews for all data sets, so applications of sentiment analysis using mixed gender data sets will over represent the opinions of women. Training on same gender data improves performance less than having additional data from both genders. Practical implications End users of sentiment analysis should be aware that its small gender biases can affect the conclusions drawn from it and apply correction factors when necessary. Users of systems that incorporate sentiment analysis should be aware that performance will vary by author gender. Developers do not need to create gender-specific algorithms unless they have more training data than their system can cope with. Originality/value This is the first demonstration of gender bias in machine learning sentiment analysis.

Publisher

Emerald

Subject

Library and Information Sciences,Computer Science Applications,Information Systems

Reference50 articles.

1. Altmetric: enriching scholarly content with article-level discussion and metrics;Learned Publishing,2013

2. Seeing without knowing: limitations of the transparency ideal and its application to algorithmic accountability;New Media & Society,2018

3. Bolukbasi, T., Chang, K.W., Zou, J.Y., Saligrama, V. and Kalai, A.T. (2016), “Man is to computer programmer as woman is to homemaker? Debiasing word embeddings”, Advances in Neural Information Processing Systems 29 (NIPS2016), Neural Information Processing Systems Foundation, Inc., Barcelona, pp. 4349-4357.

4. Burger, J.D., Henderson, J., Kim, G. and Zarrella, G. (2011), “Discriminating gender on Twitter”, Proceedings of the Conference on Empirical Methods in Natural Language Processing, Association for Computational Linguistics, pp. 1301-1309.

5. Recognizing faces across continents: the effect of within-race variations on the own-race bias in face recognition;Psychonomic Bulletin & Review,2008

Cited by 15 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Leveraging social media to examine sustainability communication of home appliance brands;Technology in Society;2024-06

2. Analyzing Biases in Popular Answer Selection Datasets on Neural-Based QA Models;Lecture Notes in Computer Science;2024

3. Technology assisted research assessment: algorithmic bias and transparency issues;Aslib Journal of Information Management;2023-10-02

4. University students as early adopters of ChatGPT: Innovation Diffusion Study;2023-03-27

5. A systematic review of socio-technical gender bias in AI algorithms;Online Information Review;2023-03-14