TECHNICAL GUIDELINES FOR ASSESSING COMPUTERIZED ADAPTIVE TESTS
Journal of Educational MeasurementPublished 1 December 1984
Bert F. Green, R. Darrell Bock, Lloyd G. Humphreys, Robert L. Linn, Mark D. Reckase
Citations383
SJR quartileQ1
SJR score1.09
SNIP1.58
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
Guidelines are proposed for evaluating a computerized adaptive test and Topics include dimensionality, measurement error, validity, estimation of item parameters, item pool characteristics and human factors.
Abstract
Guidelines are proposed for evaluating a computerized adaptive test. Topics include dimensionality, measurement error, validity, estimation of item parameters, item pool characteristics and human factors. Equating CAT and conventional tests is considered and matters of equity are addressed.
Keywords
Computer ScienceDecision Sciences
American Educational Research JournalStatistical Theories of Mental Test Scores
8,138 Citations1969William W. Rozeboom, Frederic M. Lord +2 more
Applications of Item Response Theory To Practical Testing Problems
5,024 Citations2012F. Lord
PsychometrikaMarginal Maximum Likelihood Estimation of Item Parameters: Application of an EM Algorithm
2,340 Citations1981R. Darrell Bock, Murray Aitkin
The Em procedure is shown to apply to general item-response models lacking simple sufficient statistics for ability, including models with more than one latent dimension, when computing procedures based on an EM algorithm are used.
PsychometrikaFitting a Response Model for <i>n</i> Dichotomously Scored Items
619 Citations1970R. Darrell Bock, Marcus Lieberman
British Journal of Mathematical and Statistical PsychologyThe dimensionality of tests and items
509 Citations1981Roderick P. McDonald
Journal of the American Statistical AssociationDiscrete Statistical Models With Social Science Applications.
323 Citations1983Clifford C. Clogg, Erling B. Andersen
Applied Psychological MeasurementUsing Simulation Results to Choose a Latent Trait Model
306 Citations1981Wendy M. Yen
A latent trait model goodness-of-fit statistic was defined, and its relationships to several other com monly used fit statistics were described, and some practical problems that can result from using an inappropriate model with multiple-choice tests are discussed.
Journal of the American Statistical AssociationA Bayesian Sequential Procedure for Quantal Response in the Context of Adaptive Mental Testing
284 Citations1975Roger Owen
Journal of the Royal Statistical Society Series B (Statistical Methodology)Factor Analysis for Categorical Data
194 Citations1980David J. Bartholomew
PsychometrikaBayesian Estimation in the Two-Parameter Logistic Model
167 Citations1985Hariharan Swaminathan, Janice A. Gifford
PsychometrikaThe Effect of Difficulty and Chance Success on Correlations between Items or between Tests
144 Citations1945John B. Carroll
Applied Psychological MeasurementA Broad-Range Tailored Test of Verbal Ability
121 Citations1977Frederic M. Lord
PsychometrikaSome Improved Diagnostics for Failure of the Rasch Model
111 Citations1983Ivo W. Molenaar
Applied Psychological MeasurementSome Applications of Logistic Latent Trait Models with Linear Constraints on the Parameters
78 Citations1982Gerhard Fischer, Anton K. Formann
Applied Psychological MeasurementA Use of the Information Function in Tailored Testing
78 Citations1977Fumiko Samejima
It is emphasized that the standard error of estimation should be considered as the major index of dependability, as opposed to the reliability of a test.
ETS Research Report SeriesAN EMPIRICAL STUDY OF THE BROAD RANGE TAILORED TEST OF VERBAL ABILITY
15 Citations1980Charles B. Kreitzberg, Douglas H. Jones
Applied Psychological MeasurementThe Effects of Item Calibration Sample Size and Item Pool Size on Adaptive Testing
10 Citations1981Malcolm James Ree
Evaluation Plan for the Computerized Adaptive Vocational Aptitude Battery
9 Citations1982Bert F. Green
Computer presentation, recording, and scoring of the ASVAB will improve test security and improve efficiency greatly by assessing each candidate's answers as the test progresses and posing items most appropriate for that candidate, thus avoiding items that are too easy or too hard.
History of the Armed Services Vocational Aptitude Battery (ASVAB) 1974-1980
5 Citations1980M. A. Fischl, Steven Gorman +3 more
Munich Personal RePEc Archive (Ludwig Maximilian University of Munich)The Formation of Homogeneous Item Sets When Guessing is a Factor in Item Responses.
4 Citations1981Mark D. Reckase
Munich Personal RePEc Archive (Ludwig Maximilian University of Munich)Ability Estimation and Item Calibration Using the One and Three Parameter Logistic Models: A Comparative Study.
4 Citations1977Mark D. Reckase
The literature on latent trait calibration procedures was reviewed to determine the methods available to calibrate dichotomous items for tailored testing applications and the most promising techniques using the one- and three-parameter logistic models were selected.
