| Use of Data Mining Methods to Detect Test Fraud |
6 |
| A New Person-Fit Statistic for the Lognormal Model for Response Times |
6 |
| Measures of Agreement to Assess Attribute-Level Classification Accuracy and Consistency for Cognitive Diagnostic Assessments |
5 |
| Standard Errors of IRT Parameter Scale Transformation Coefficients: Comparison of Bootstrap Method, Delta Method, and Multiple Imputation Method |
5 |
| Modeling Basic Writing Processes From Keystroke Logs |
5 |
| A Method for Detecting Regression of Hard and Easy Item Angoff Ratings |
4 |
| Modeling Response Styles in Cross-Country Self-Reports: An Application of a Multilevel Multidimensional Nominal Response Model |
4 |
| Development of Information Functions and Indices for the GGUM-RANK Multidimensional Forced Choice IRT Model |
4 |
| Gathering and Evaluating Validity Evidence: The Generalized Assessment Alignment Tool |
4 |
| Conceptualizing Rater Judgments and Rating Processes for Rater-Mediated Assessments |
3 |
| Calculating Conditional Reliability for Dynamic Measurement Model Capacity Estimates |
3 |
| Bayesian Model Selection Methods for Multilevel IRT Models: A Comparison of Five DIC-Based Indices |
3 |
| Exploring the Influence of Range Restrictions on Connectivity in Sparse Assessment Networks: An Illustration and Exploration Within the Context of Classroom Observations |
3 |
| A Top-Down Approach to Designing the Computerized Adaptive Multistage Test |
3 |
| Use of Adjustment by Minimum Discriminant Information in Linking Constructed-Response Test Scores in the Absence of Common Items |
3 |
| Parameter Invariance and Skill Attribute Continuity in the DINA Model |
2 |
| Bias and Bias Correction Method for Nonproportional Abilities Requirement (NPAR) Tests |
2 |
| The Effects of Incomplete Rating Designs in Combination With Rater Effects |
2 |
| A Comparison of Strategies for Smoothing Parameter Selection for Mixed-Format Tests Under the Random Groups Design |
2 |
| A New Facets Model for Rater's Centrality/Extremity Response Style |
2 |
| Evaluating Intervention Effects in a Diagnostic Classification Model Framework |
2 |
| Cross-Country Heterogeneity in Students' Reporting Behavior: The Use of the Anchoring Vignette Method |
2 |
| An Alternative to the 3PL: Using Asymmetric Item Characteristic Curves to Address Guessing Effects |
1 |
| Detecting Nonadditivity in Single-Facet Generalizability Theory Applications: Tukey's Test |
1 |
| On-the-Fly Constraint-Controlled Assembly Methods for Multistage Adaptive Testing for Cognitive Diagnosis |
1 |
| Using Hierarchical Logistic Regression to Study DIF and DIF Variance in Multilevel Data |
1 |
| Pedagogical Considerations for Examining Rater Variability in Rater-Mediated Assessments: A Three-Model Framework |
1 |
| Assessing and Validating Effects of a Data-Based Decision-Making Intervention on Student Growth for Mathematics and Spelling |
1 |
| Exploring How to Model Formative Assessment Trajectories of Posing-Pausing-Probing Practices: Toward a Teacher Learning Progressions Framework for the Study of Novice Teachers |
1 |
| Students' Interpretation of Formative Assessment Feedback: Three Claims for Why We Know So Little About Something So Important |
1 |
| Examining Differential Rater Functioning Using a Between-Subgroup Outfit Approach |
1 |
| Scale Alignment in Between-Item Multidimensional Rasch Models |
1 |
| Item Response Models for Multiple Attempts With Incomplete Data |
1 |
| Computerized Adaptive Testing in Early Education: Exploring the Impact of Item Position Effects on Ability Estimation |
1 |
| Subjective Priors for Item Response Models: Application of Elicitation by Design |
1 |
| Examining the Dual Purpose Use of Student Learning Objectives for Classroom Assessment and Teacher Evaluation |
1 |
| A Comparison of Experimental and Observational Approaches to Assessing the Effects of Time Constraints in a Medical Licensing Examination |
1 |
| A General Framework for the Validation of Embedded Formative Assessment |
1 |
| Lord's Wald Test for Detecting DIF in Multidimensional IRT Models: A Comparison of Two Estimation Approaches |
0 |
| Classroom Assessment and Large-Scale Psychometrics: Shall the Twain Meet? (A Conversation With Margaret Heritage and Neal Kingston) |
0 |
| Controlling Bias in Both Constructed Response and Multiple-Choice Items When Analyzed With the Dichotomous Rasch Model |
0 |
| The Impact of Multidimensionality on Extraction of Latent Classes in Mixture Rasch Models |
0 |
| A Comparison of Procedures for Estimating Person Reliability Parameters in the Graded Response Model |
0 |
| Effectiveness of Equating at the Passing Score for Exams With Small Sample Sizes |
0 |
| Comparing Academic Readiness Requirements for Different Postsecondary Pathways: What Admissions Tests Tell Us |
0 |
| Modeling Partial Knowledge on Multiple-Choice Items Using Elimination Testing |
0 |
| Examining Psychometric Properties and Level Classification of the van Hiele Geometry Test Using CTT and CDM Frameworks |
0 |
| A New Interpretation of Augmented Subscores and Their Added Value in Terms of Parallel Forms |
0 |
| Routing Strategies and Optimizing Design for Multistage Testing in International Large-Scale Assessments |
0 |
| Efficiency of Targeted Multistage Calibration Designs Under Practical Constraints: A Simulation Study |
0 |