State of the psychometric methods: patient-reported outcome measure development and refinement using item response theory

被引:57
作者
Stover, Angela M. [1 ,2 ]
McLeod, Lori D. [3 ]
Langer, Michelle M. [2 ,4 ,5 ]
Chen, Wen-Hung [3 ]
Reeve, Bryce B. [1 ,6 ]
机构
[1] Univ N Carolina, Dept Hlth Policy & Management, 1101-G McGavran Greenberg Hall CB 7411, Chapel Hill, NC 27599 USA
[2] Univ N Carolina, Lineberger Comprehens Canc Ctr, Sch Med, 101 Manning Dr, Chapel Hill, NC 27599 USA
[3] RTI Hlth Solut, 3040 Cornwallis Rd, Res Triangle Pk, NC 27709 USA
[4] Northwestern Univ, Med Social Sci, 625 N Michigan Ave Suite 2700, Chicago, IL 60611 USA
[5] Northwestern Univ, Feinberg Sch Med, 625 N Michigan Ave Suite 2700, Chicago, IL 60611 USA
[6] Duke Univ, Sch Med, Dept Populat Hlth Sci & Pediat, Ctr Hlth Measurement, 2200 West Main St,Suite 720A, Durham, NC 27707 USA
关键词
Item response theory; Scale construction; Scale evaluation; Measurement; PROMIS (R); GOODNESS-OF-FIT; INFORMATION-SYSTEM PROMIS(R); INSTRUMENT DEVELOPMENT; LIMITED-INFORMATION; LATENT ABILITY; IRT MODEL; DEPRESSION; IMPACT; ORGANIZATION; VALIDATION;
D O I
10.1186/s41687-019-0130-5
中图分类号
R19 [保健组织与事业(卫生事业管理)];
学科分类号
摘要
Background: This paper is part of a series comparing different psychometric approaches to evaluate patient-reported outcome (PRO) measures using the same items and dataset. We provide an overview and example application to demonstrate 1) using item response theory (IRT) to identify poor and well performing items; 2) testing if items perform differently based on demographic characteristics (differential item functioning, DIF); and 3) balancing IRT and content validity considerations to select items for short forms. Methods: Model fit, local dependence, and DIF were examined for 51 items initially considered for the Patient-Reported Outcomes Measurement Information System (R) (PROMIS (R)) Depression item bank. Samejima's graded response model was used to examine how well each item measured severity levels of depression and how well it distinguished between individuals with high and low levels of depression. Two short forms were constructed based on psychometric properties and consensus discussions with instrument developers, including psychometricians and content experts. Calibrations presented here are for didactic purposes and are not intended to replace official PROMIS parameters or to be used for research. Results: Of the 51 depression items, 14 exhibited local dependence, 3 exhibited DIF for gender, and 9 exhibited misfit, and these items were removed from consideration for short forms. Short form 1 prioritized content, and thus items were chosen to meet DSM-V criteria rather than being discarded for lower discrimination parameters. Short form 2 prioritized well performing items, and thus fewer DSM-V criteria were satisfied. Short forms 1-2 performed similarly for model fit statistics, but short form 2 provided greater item precision. Conclusions: IRT is a family of flexible models providing item- and scale-level information, making it a powerful tool for scale construction and refinement. Strengths of IRT models include placing respondents and items on the same metric, testing DIF across demographic or clinical subgroups, and facilitating creation of targeted short forms. Limitations include large sample sizes to obtain stable item parameters, and necessary familiarity with measurement methods to interpret results. Combining psychometric data with stakeholder input (including people with lived experiences of the health condition and clinicians) is highly recommended for scale development and evaluation.
引用
收藏
页数:16
相关论文
共 50 条
[31]   RespOnse Shift ALgorithm in Item response theory (ROSALI) for response shift detection with missing data in longitudinal patient-reported outcome studies [J].
Guilleux, Alice ;
Blanchin, Myriam ;
Vanier, Antoine ;
Guillemin, Francis ;
Falissard, Bruno ;
Schwartz, Carolyn E. ;
Hardouin, Jean-Benoit ;
Sebille, Veronique .
QUALITY OF LIFE RESEARCH, 2015, 24 (03) :553-564
[32]   Stroke Social Network Scale: development and psychometric evaluation of a new patient-reported measure [J].
Northcott, Sarah ;
Hilari, Katerina .
CLINICAL REHABILITATION, 2013, 27 (09) :823-833
[33]   Using item response theory improved responsiveness of patient-reported outcomes measures in carpal tunnel syndrome [J].
Lyren, Per-Erik ;
Atroshi, Isam .
JOURNAL OF CLINICAL EPIDEMIOLOGY, 2012, 65 (03) :325-334
[34]   Development of a Conceptual Framework and Calibrated Item Banks to Measure Patient-Reported Dyspnea Severity and Related Functional Limitations [J].
Choi, Seung W. ;
Victorson, David E. ;
Yount, Susan ;
Anton, Susan ;
Cella, David .
VALUE IN HEALTH, 2011, 14 (02) :291-306
[35]   Development of a new patient-reported outcome measure for Dupuytren disease: A study protocol [J].
Eckerdal, David ;
Lyren, Per-Erik ;
Mceachan, Jane ;
Lauritzson, Anna ;
Nordenskjold, Jesper ;
Atroshi, Isam .
HEALTH INFORMATICS JOURNAL, 2024, 30 (04)
[36]   The Concerns About Pain (CAP) Scale: A Patient-Reported Outcome Measure of Pain Catastrophizing [J].
Amtmann, Dagmar ;
Bamer, Alyssa M. ;
Liljenquist, Kendra S. ;
Cowan, Penney ;
Salem, Rana ;
Turk, Dennis C. ;
Jensen, Mark P. .
JOURNAL OF PAIN, 2020, 21 (11-12) :1198-1211
[37]   Item Banking: A Generational Change in Patient-Reported Outcome Measurement [J].
Pesudovs, Konrad .
OPTOMETRY AND VISION SCIENCE, 2010, 87 (04) :285-293
[38]   Mental Pain as a Transdiagnostic Patient-Reported Outcome Measure [J].
Fava, Giovanni A. ;
Tomba, Elena ;
Brakemeier, Eva-Lotta ;
Carrozzino, Danilo ;
Cosci, Fiammetta ;
Eory, Ajandek ;
Leonardi, Tommaso ;
Schamong, Isabel ;
Guidi, Jenny .
PSYCHOTHERAPY AND PSYCHOSOMATICS, 2019, 88 (06) :341-349
[39]   A brief measure of problematic smartphone use among high school students: Psychometric assessment using item response theory [J].
Donaldson, Scott I. ;
Strong, David ;
Zhu, Shu-Hong .
COMPUTERS IN HUMAN BEHAVIOR REPORTS, 2021, 3
[40]   Psychometric Properties of Patient-Reported Outcome Measures for Periacetabular Osteotomy [J].
Wasko, Marcin K. ;
Yanik, Elizabeth L. ;
Pascual-Garrido, Cecilia ;
Clohisy, John C. .
JOURNAL OF BONE AND JOINT SURGERY-AMERICAN VOLUME, 2019, 101 (06)