Measuring test-retest reliability (TRR) of AMSTAR provides moderate to perfect agreement – a contribution to the discussion of the importance of TRR in relation to the psychometric properties of assessment tools-Reference-Cited by-同舟云学术

Measuring test-retest reliability (TRR) of AMSTAR provides moderate to perfect agreement – a contribution to the discussion of the importance of TRR in relation to the psychometric properties of assessment tools

Published:2021-03-11 Issue:1 Volume:21 Page:
ISSN:1471-2288
Container-title:BMC Medical Research Methodology
language:en
Short-container-title:BMC Med Res Methodol

Author:

Bühn Stefanie,Ober Peggy,Mathes Tim,Wegewitz Uta,Jacobs Anja,Pieper Dawid

Abstract

Abstract Background Systematic Reviews (SRs) can build the groundwork for evidence-based health care decision-making. A sound methodological quality of SRs is crucial. AMSTAR (A Measurement Tool to Assess Systematic Reviews) is a widely used tool developed to assess the methodological quality of SRs of randomized controlled trials (RCTs). Research shows that AMSTAR seems to be valid and reliable in terms of interrater reliability (IRR), but the test retest reliability (TRR) of AMSTAR has never been investigated. In our study we investigated the TRR of AMSTAR to evaluate the importance of its measurement and contribute to the discussion of the measurement properties of AMSTAR and other quality assessment tools. Methods Seven raters at three institutions independently assessed the methodological quality of SRs in the field of occupational health with AMSTAR. Between the first and second ratings was a timespan of approximately two years. Answers were dichotomized, and we calculated the TRR of all raters and AMSTAR items using Gwet’s AC1 coefficient. To investigate the impact of variation in the ratings over time, we obtained summary scores for each review. Results AMSTAR item 4 (Was the status of publication used as an inclusion criterion?) provided the lowest median TRR of 0.53 (moderate agreement). Perfect agreement of all reviewers was detected for AMSTAR-item 1 with a Gwet’s AC1 of 1, which represented perfect agreement. The median TRR of the single raters varied between 0.69 (substantial agreement) and 0.89 (almost perfect agreement). Variation of two or more points in yes-scored AMSTAR items was observed in 65% (73/112) of all assessments. Conclusions The high variation between the first and second AMSTAR ratings suggests that consideration of the TRR is important when evaluating the psychometric properties of AMSTAR.. However, more evidence is needed to investigate this neglected issue of measurement properties. Our results may initiate discussion of the importance of considering the TRR of assessment tools. A further examination of the TRR of AMSTAR, as well as other recently established rating tools such as AMSTAR 2 and ROBIS (Risk Of Bias In Systematic reviews), would be useful.

Funder

Private Universität Witten/Herdecke gGmbH

Publisher

Springer Science and Business Media LLC

Subject

Health Informatics,Epidemiology

Link

http://link.springer.com/content/pdf/10.1186/s12874-021-01231-y.pdf

Reference31 articles.

1. Ioannidis JP. The mass production of redundant, misleading, and conflicted systematic reviews and meta-analyses. Milbank Q. 2016;94(3):485–514.

2. Bastian H, Glasziou P, Chalmers I. Seventy-five trials and eleven systematic reviews a day: how will we ever keep up? PLoS Med. 2010;7(9):e1000326.

3. Pollock M FR, Becker LA, Pieper D, Hartling L.. Chapter V: Overviews of Reviews [Available from: www.training.cochrane.org/handbook. Accessed Nov 2020.

4. Gwet KL. Handbook of inter-rater reliability: the definitive guide to measuring the extent of agreement among raters: advanced analytics. United States of America: LLC; 2014.

5. Shea BJ, Hamel C, Wells GA, Bouter LM, Kristjansson E, Grimshaw J, et al. AMSTAR is a reliable and valid measurement tool to assess the methodological quality of systematic reviews. J Clin Epidemiol. 2009;62(10):1013–20.

Cited by 6 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. The Finishing Space Value for Shooting Decision-Making in High-Performance Football;Sports;2024-07-30

2. Evaluation of methodological and reporting quality of systematic reviews on conservative non-pharmacological musculoskeletal pain management in children and adolescents: A methodological analysis;Musculoskeletal Science and Practice;2024-02

3. Quality assessment of systematic reviews with meta-analysis in undergraduate nursing education;Nurse Education Today;2023-07

4. Efficacy and safety of novel oral anticoagulants for the treatment of cancer-associated venous thromboembolism: protocol for an umbrella review of systematic reviews and meta-analyses;BMJ Open;2023-04

5. Efficacy of Therapist Supported Interventions from the Neonatal Intensive Care Unit to Home;Clinics in Perinatology;2023-03