Validation Methods for Aggregate-Level Test Scale Linking: A Case Study Mapping School District Test Score Distributions to a Common Scale-Reference-Cited by-同舟云学术

Validation Methods for Aggregate-Level Test Scale Linking: A Case Study Mapping School District Test Score Distributions to a Common Scale

Published:2019-10-08 Issue:2 Volume:46 Page:138-167
ISSN:1076-9986
Container-title:Journal of Educational and Behavioral Statistics
language:en
Short-container-title:Journal of Educational and Behavioral Statistics

Author:

Reardon Sean F.,Kalogrides Demetra¹,Ho Andrew D.²

Affiliation:

1. Stanford University

2. Harvard Graduate School of Education

Abstract

Linking score scales across different tests is considered speculative and fraught, even at the aggregate level. We introduce and illustrate validation methods for aggregate linkages, using the challenge of linking U.S. school district average test scores across states as a motivating example. We show that aggregate linkages can be validated both directly and indirectly under certain conditions such as when the scores for at least some target units (districts) are available on a common test (e.g., the National Assessment of Educational Progress). We introduce precision-adjusted random effects models to estimate linking error, for populations and for subpopulations, for averages and for progress over time. These models allow us to distinguish linking error from sampling variability and illustrate how linking error plays a larger role in aggregates with smaller sample sizes. Assuming that target districts generalize to the full population of districts, we can show that standard errors for district means are generally less than .2 standard deviation units, leading to reliabilities above .7 for roughly 90% of districts. We also show how sources of imprecision and linking error contribute to both within- and between-state district comparisons within versus between states. This approach is applicable whenever the essential counterfactual question—“what would means/variance/progress for the aggregate units be, had students taken the other test?”—can be answered directly for at least some of the units.

Funder

Institute of Education Sciences

Spencer Foundation

William T. Grant Foundation

Publisher

American Educational Research Association (AERA)

Subject

Social Sciences (miscellaneous),Education

Link

http://journals.sagepub.com/doi/pdf/10.3102/1076998619874089

Cited by 23 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Generally Applicable Variance Estimation Methods for Common-Population Linking;Journal of Educational and Behavioral Statistics;2024-08-08

2. Closing Reporting Gaps: A Comparison of Methods for Estimating Unreported Subgroup Achievement on NAEP;Measurement: Interdisciplinary Research and Perspectives;2024-01-02

3. What Impacts Can We Expect from School Spending Policy? Evidence from Evaluations in the United States;American Economic Journal: Applied Economics;2024-01-01

4. The distribution of child physicians and early academic achievement;Health Services Research;2023-06-07

5. Free and reduced-price meal enrollment does not measure student poverty: Evidence and policy significance;Economics of Education Review;2023-06