ERIC - Search Results

Publication Date

In 2026	0
Since 2025	0
Since 2022 (last 5 years)	0
Since 2017 (last 10 years)	0
Since 2007 (last 20 years)	7

Descriptor

Accuracy	8
Statistical Analysis	8
Equated Scores	5
Sample Size	5
Computation	4
Statistical Significance	3
Comparative Analysis	2
Differences	2
Error of Measurement	2
Models	2
Regression (Statistics)	2
Scores	2
Simulation	2
Adaptive Testing	1
Bayesian Statistics	1
Change	1
Data	1
Data Analysis	1
Difficulty Level	1
Effect Size	1
Goodness of Fit	1
Item Response Theory	1
Measures (Individuals)	1
Selection	1
Statistical Bias	1
More ▼

Source

ETS Research Report Series	5
Educational Testing Service	1
Educational and Psychological…	1
Journal of Educational…	1

Author

Moses, Tim	8
Holland, Paul	3
Kim, Sooyeon	2
Dorans, Neil	1
Miao, Jing	1
Yoo, Hanwook	1

Publication Type

Journal Articles	7
Reports - Research	6
Reports - Evaluative	2

Education Level

Audience

Location

Laws, Policies, & Programs

Assessments and Surveys

What Works Clearinghouse Rating

Showing all 8 results Save | Export

A Comparison of IRT Proficiency Estimation Methods under Adaptive Multistage Testing

Peer reviewed

Direct link

Kim, Sooyeon; Moses, Tim; Yoo, Hanwook – Journal of Educational Measurement, 2015

This inquiry is an investigation of item response theory (IRT) proficiency estimators' accuracy under multistage testing (MST). We chose a two-stage MST design that includes four modules (one at Stage 1, three at Stage 2) and three difficulty paths (low, middle, high). We assembled various two-stage MST panels (i.e., forms) by manipulating two…

Descriptors: Comparative Analysis, Item Response Theory, Computation, Accuracy

Evaluating Ranking Strategies in Assessing Change when the Measures Differ across Time

Peer reviewed

Direct link

Moses, Tim; Kim, Sooyeon – Educational and Psychological Measurement, 2012

In this study, a ranking strategy was evaluated for comparing subgroups' change using identical, equated, and nonidentical measures. Four empirical data sets were evaluated, each of which contained examinees' scores on two occasions, where the two occasions' scores were obtained on a single identical measure, on two equated tests, and on two…

Descriptors: Testing, Change, Scores, Measures (Individuals)

Alternative Loglinear Smoothing Models and Their Effect on Equating Function Accuracy. Research Report. ETS RR-09-48

Peer reviewed
PDF on ERIC

Download full text

Moses, Tim; Holland, Paul – ETS Research Report Series, 2009

This simulation study evaluated the potential of alternative loglinear smoothing strategies for improving equipercentile equating function accuracy. These alternative strategies use cues from the sample data to make automatable and efficient improvements to model fit, either through the use of indicator functions for fitting large residuals or by…

Descriptors: Accuracy, Equated Scores, Statistical Analysis, Models

A Comparison of Methods for Estimating Conditional Item Score Differences in Differential Item Functioning (DIF) Assessments. Research Report. ETS RR-10-15

Download full text

Moses, Tim; Miao, Jing; Dorans, Neil – Educational Testing Service, 2010

This study compared the accuracies of four differential item functioning (DIF) estimation methods, where each method makes use of only one of the following: raw data, logistic regression, loglinear models, or kernel smoothing. The major focus was on the estimation strategies' potential for estimating score-level, conditional DIF. A secondary focus…

Descriptors: Test Bias, Statistical Analysis, Computation, Scores

The Influence of Strategies for Selecting Loglinear Smoothing Models on Equating Functions. Research Report. ETS RR-08-25

Peer reviewed
PDF on ERIC

Download full text

Moses, Tim; Holland, Paul – ETS Research Report Series, 2008

This study addressed 2 issues of using loglinear models for smoothing univariate test score distributions and for enhancing the stability of equipercentile equating functions. One issue was a comparative assessment of several statistical strategies that have been proposed for selecting 1 from several competing model parameterizations. Another…

Descriptors: Equated Scores, Selection, Models, Statistical Analysis

An Evaluation of Statistical Strategies for Making Equating Function Selections. Research Report. ETS RR-08-60

Peer reviewed
PDF on ERIC

Download full text

Moses, Tim – ETS Research Report Series, 2008

Nine statistical strategies for selecting equating functions in an equivalent groups design were evaluated. The strategies of interest were likelihood ratio chi-square tests, regression tests, Kolmogorov-Smirnov tests, and significance tests for equated score differences. The most accurate strategies in the study were the likelihood ratio tests…

Descriptors: Equated Scores, Statistical Analysis, Statistical Significance, Regression (Statistics)

Kernel and Traditional Equipercentile Equating with Degrees of Presmoothing. Research Report. ETS RR-07-15

Peer reviewed
PDF on ERIC

Download full text

Moses, Tim; Holland, Paul – ETS Research Report Series, 2007

The purpose of this study was to empirically evaluate the impact of loglinear presmoothing accuracy on equating bias and variability across chained and post-stratification equating methods, kernel and percentile-rank continuization methods, and sample sizes. The results of evaluating presmoothing on equating accuracy generally agreed with those of…

Descriptors: Equated Scores, Statistical Analysis, Accuracy, Sample Size

Using the Kernel Method of Test Equating for Estimating the Standard Errors of Population Invariance Measures. Research Report. ETS RR-06-20

Peer reviewed
PDF on ERIC

Download full text

Moses, Tim – ETS Research Report Series, 2006

Population invariance is an important requirement of test equating. An equating function is said to be population invariant when the choice of (sub)population used to compute the equating function does not matter. In recent studies, the extent to which equating functions are population invariant is typically addressed in terms of practical…

Descriptors: Equated Scores, Computation, Error of Measurement, Statistical Analysis