Patient Health Questionnaire-9 Item Pairing Predictiveness for Prescreening Depressive Symptomatology: Machine Learning Analysis (Preprint)

Author:

Glavin DarraghORCID,Grua Eoin MartinoORCID,Nakamura Carina AkemiORCID,Scazufca MarciaORCID,Ribeiro dos Santos EdinilzaORCID,Wong Gloria H YORCID,Hollingworth WilliamORCID,Peters Tim JORCID,Araya RicardoORCID,Van de Ven PepijnORCID

Abstract

BACKGROUND

<i>Anhedonia</i> and <i>depressed mood</i> are considered the cardinal symptoms of major depressive disorder. These are the first 2 items of the Patient Health Questionnaire (PHQ)–9 and comprise the ultrabrief PHQ-2 used for prescreening depressive symptomatology. The prescreening performance of alternative PHQ-9 item pairings is rarely compared with that of the PHQ-2.

OBJECTIVE

This study aims to use machine learning (ML) with the PHQ-9 items to identify and validate the most predictive 2-item depressive symptomatology ultrabrief questionnaire and to test the generalizability of the best pairings found on the primary data set, with 6 external data sets from different populations to validate their use as prescreening instruments.

METHODS

All 36 possible PHQ-9 item pairings (each yielding scores of 0-6) were investigated using ML-based methods with logistic regression models. Their performances were evaluated based on the classification of depressive symptomatology, defined as PHQ-9 scores ≥10. This gave each pairing an equal opportunity and avoided any bias in item pairing selection.

RESULTS

The ML-based PHQ-9 items 2 and 4 (phq2&amp;4), the <i>depressed mood</i> and <i>low-energy</i> item pairing, and PHQ-9 items 2 and 8 (phq2&amp;8), the <i>depressed mood</i> and <i>psychomotor retardation or agitation</i> item pairing, were found to be the best on the primary data set training split. They generalized well on the primary data set test split with area under the curves (AUCs) of 0.954 and 0.946, respectively, compared with an AUC of 0.942 for the PHQ-2. The phq2&amp;4 had a higher AUC than the PHQ-2 on all 6 external data sets, and the phq2&amp;8 had a higher AUC than the PHQ-2 on 3 data sets. The phq2&amp;4 had the highest Youden index (an unweighted average of sensitivity and specificity) on 2 external data sets, and the phq2&amp;8 had the highest Youden index on another 2. The PHQ-2≥2 cutoff also had the highest Youden index on 2 external data sets, joint highest with the phq2&amp;4 on 1, but its performance fluctuated the most. The PHQ-2≥3 cutoff had the highest Youden index on 1 external data set. The sensitivity and specificity achieved by the phq2&amp;4 and phq2&amp;8 were more evenly balanced than the PHQ-2≥2 and ≥3 cutoffs.

CONCLUSIONS

The PHQ-2 did not prove to be a more effective prescreening instrument when compared with other PHQ-9 item pairings. Evaluating all item pairings showed that, compared with alternative partner items, the <i>anhedonia</i> item underperformed alongside the <i>depressed mood</i> item. This suggests that the inclusion of <i>anhedonia</i> as a core symptom of depression and its presence in ultrabrief questionnaires may be incompatible with the empirical evidence. The use of the PHQ-2 to prescreen for depressive symptomatology could result in a greater number of misclassifications than alternative item pairings.

Publisher

JMIR Publications Inc.

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3