Educational and Psychological Measurement

Papers
(The TQCC of Educational and Psychological Measurement is 4. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Assessing the Properties and Functioning of Model-Based Sum Scores in Multidimensional Measures With Local Item Dependencies: A Comprehensive Proposal385
Using Deep Reinforcement Learning to Decide Test Length44
Iterative Item Selection of Neighborhood Clusters: A Nonparametric and Non-IRT Method for Generating Miniature Computer Adaptive Questionnaires33
Functional Approaches for Modeling Unfolding Data31
Model Specification Searches in Structural Equation Modeling Using Bee Swarm Optimization23
Assessing the Unconditional and Conditional External Validity of Noncognitive Test Scores: A Unifying Model-Based Proposal18
Collapsing Sparse Responses in Likert-Type Scale Data: Advantages and Disadvantages for Model Fit in CFA18
An Illustration of an IRTree Model for Disengagement17
Detecting Preknowledge Cheating via Innovative Measures: A Mixture Hierarchical Model for Jointly Modeling Item Responses, Response Times, and Visual Fixation Counts17
Using Item Scores and Response Times to Detect Item Compromise in Computerized Adaptive Testing15
Detecting Differential Item Functioning Using Response Time15
Generalized Mantel–Haenszel Estimators for Simultaneous Differential Item Functioning Tests12
Optimal Number of Replications for Obtaining Stable Dynamic Fit Index Cutoffs12
Item Parameter Recovery: Sensitivity to Prior Distribution11
Reconceptualizing Scoring Reliability Through Linguistic Similarity11
On the Benefits of Using Maximal Reliability in Educational and Behavioral Research11
Multimodal Test Item Parameter Prediction From Text, Images, and Metadata: Fusing Together AI Vision and Language Models11
Improving the Use of Parallel Analysis by Accounting for Sampling Variability of the Observed Correlation Matrix10
An Explanatory Multidimensional Random Item Effects Rating Scale Model10
How to Improve the Regression Factor Score Predictor When Individuals Have Different Factor Loadings9
Examining the Dynamic of Clustering Effects in Multilevel Designs: A Latent Variable Method Application9
Examination of ChatGPT’s Performance as a Data Analysis Tool9
An Omega-Hierarchical Extension Index for Second-Order Constructs With Hierarchical Measuring Instruments9
Rotation Local Solutions in Multidimensional Item Response Theory Models9
Assessing the Speed–Accuracy Tradeoff in Psychological Testing Using Experimental Manipulations9
What Affects the Quality of Score Transformations? Potential Issues in True-Score Equating Using the Partial Credit Model9
Measuring Unipolar Traits With Continuous Response Items: Some Methodological and Substantive Developments8
Integrating Ensemble Clustering and Text Embeddings for Estimating the Factor Loadings of Self-Report Scales8
From Linear Geometry to Nonlinear and Information-Geometric Settings in Test Theory: Bregman Projections as a Unifying Framework8
Separation of Traits and Extreme Response Style in IRTree Models: The Role of Mimicry Effects for the Meaningful Interpretation of Estimates8
The Aggregated Latent Profile Index: Measuring Person Profile Differentiation Within a Bootstrap-Validated Latent Profile Space8
Using Multiple Imputation to Account for the Uncertainty Due to Missing Data in the Context of Factor Retention7
Obtaining a Bayesian Estimate of Coefficient Alpha Using a Posterior Normal Distribution7
Agreement Lambda for Weighted Disagreement With Ordinal Scales: Correction for Category Prevalence6
On the Complex Sources of Differential Item Functioning: A Comparison of Three Methods6
A Bayesian General Model to Account for Individual Differences in Operation-Specific Learning Within a Test6
Reliability of Difference Scores Obtained From Nested Data Within a Multivariate Generalizability Theory Framework5
Detecting Rating Scale Malfunctioning With the Partial Credit Model and Generalized Partial Credit Model5
Differential Item Functioning Effect Size Use for Validity Information5
Linear and Nonlinear Indices of Score Accuracy and Item Effectiveness for Measures That Contain Locally Dependent Items5
Are Speeded Tests Unfair? Modeling the Impact of Time Limits on the Gender Gap in Mathematics5
Discriminating Between Attribute, Item-Position, and Wording Effects by the Congeneric and Tau-Equivalent Confirmatory Factor Analysis Models5
The One-Parameter Logistic Model Can Be True With Zero Probability for a Unidimensional Measuring Instrument: How One Could Go Wrong Removing Items Not Satisfying the Model5
Evaluating Model Fit of Measurement Models in Confirmatory Factor Analysis5
Examining the Instructional Sensitivity of Constructed-Response Achievement Test Item Scores5
Evaluating the Performance of a Regularized Differential Item Functioning Method for Testlet-Based Polytomous Items4
Reducing Calibration Bias for Person Fit Assessment by Mixture Model Expansion4
An Evaluation of Fit Indices Used in Model Selection of Dichotomous Mixture IRT Models4
Investigating Heterogeneity in Response Strategies: A Mixture Multidimensional IRTree Approach4
Impacts of DIF Item Balance and Effect Size Incorporation With the Rasch Tree4
Equidistant Response Options on Likert-Type Instruments: Testing the Interval Scaling Assumption Using Mplus4
An Item Response Theory Model for Incorporating Response Times in Forced-Choice Measures4
Historical Measurement Information Can Be Used to Improve Estimation of Structural Parameters in Structural Equation Models With Small Samples4
Overestimation of Internal Consistency by Coefficient Omega in Data Giving Rise to a Centroid-Like Factor Solution4
Identification and Diagnosis of Misreporting in Surveys4
Is Effort Moderated Scoring Robust to Multidimensional Rapid Guessing?4
Modeling Misspecification as a Parameter in Bayesian Structural Equation Models4
0.1661741733551