Measurement & scaling
Social-emotional learning competencies are important for student success, but are they stable over time? This study explores this question and the implications for teachers and schools.
By: James Soland, Megan Kuhfeld, Emily Wolk, Sharon Bi
Topics: Measurement & scaling, Social-emotional learning, Student growth & accountability policies
Predicting time to reclassification for English learners: A joint modeling approach
The development of academic English proficiency and the time it takes to reclassify to fluent English proficient status are key issues in English learner (EL) policy. This article develops a shared random effects model (SREM) to estimate English proficiency development and time to reclassification simultaneously, treating student-specific random effects as latent covariates in the time to reclassification model.
By: Tyler Matta, James Soland
This paper briefly discusses the trade-offs involved in making such a transition, and then focuses on a relatively unexplored benefit of computer-based tests ā the control of construct-irrelevant factors that can threaten test score validity.
By: Steven Wise
Topics: Measurement & scaling, Innovations in reporting & assessment, School & test engagement
When computer-based tests are used, disengagement can be detected through occurrences of rapid-guessing behavior. This empirical study investigated the impact of a new effort monitoring feature that can detect rapid guessing, as it occurs, and notify proctors that a test taker has become disengaged.
By: Steven Wise, Megan Kuhfeld, James Soland
Topics: Measurement & scaling, Innovations in reporting & assessment, School & test engagement
Identifying disengaged survey responses: New evidence using response time metadata
In this study, we condition results from a variety of detection methods used to identify disengaged survey responses on response times. We then show how this conditional approach may be useful in identifying where to set response time thresholds for survey items, as well as in avoiding misclassification when using other detection methods.
By: James Soland, Steven Wise, Lingyun Gao
Robust IRT scaling: Considerations in constructing item bank from tests across years
This study investigates the impact of three different IRT scaling and equating methods in building an item bank of tests from 23 years of a national licensure exam . The study focuses on several key psychometric issues including scaledriftandequatingerrors.
By: Jungnam Kim, Dong-In Kim, Furong Gao
Topics: Measurement & scaling, Computer adaptive testing, Item response theory
This article addresses the issue by estimating teacher value added, then applying extremely mild nonlinear transformations to the original scale and re-estimating the value added. Although by definition at most one of these scales can be equal-interval, all are treated as if interval-scaled when estimating value added.
By: James Soland
Topics: Measurement & scaling, Student growth & accountability policies