Schedule At A Glance
Monday, 28 September
- Morning: Pre-Conference Workshops*
- Afternoon: Welcome Plenary and Paper Sessions
- Evening: Welcome Reception and Poster Session
| ROOM 1 | ROOM 2 | ROOM 3 | |
|---|---|---|---|
| 8:00 - 10:00 | Workshop 1 | Workshop 2 | Workshop 3 |
| 10:00 - 10:30 | COFFEE BREAK | ||
| 10:30 - 12:00 | Workshop 1 | Workshop 2 | Workshop 3 |
| 12:00 - 14:00 | LUNCH | ||
| 14:00 - 14:30 | Opening ceremony | ||
| 14:30 - 14:45 | Introduction of board members and Welcome to incoming president Chun Wang | ||
| 14:45 - 15:45 | Keynote 1: Wim van der Linden (Presidential address) Simplifying large-scale educational assessments through Bayesian adaptive testing |
||
| 15:45 - 16:15 | COFFEE BREAK | ||
| 16:15 - 17:15 | Keynote 2: Andreas Frey Selecting Empirical Prior Distributions for the Ability Parameters in Large-Scale Educational Assessments |
||
| 17:15 - 18:00 | Awards and Travel Grants | ||
| 18:00 - 20:00 | Poster session (10 works) and Welcome Reception | ||
Tuesday, 29 September
- Morning: Plenary and Paper Sessions
- Afternoon: Plenary and Paper Sessions
- Evening: Conference Dinner — Concert, Tour and Dinner at Catedral da Sé*
| ROOM 1 | ROOM 2 | ROOM 3 | |
|---|---|---|---|
| 8:00 - 9:40 | Symposia: Advancing Adaptive Measurement: Dr. David Weiss's IRT and CAT Lab Innovations in CAT Methodology and Adaptive Measures of Change | Individual presentations: Applications of CAT (Education and Language) | Individual presentations: Other AI Topics in Psychometrics |
| Change Pattern Detection in Adaptive Measure of Individual Change (Raj Wahlquist) | Development and Evaluation of an IRT-Based Computerized Adaptive Testing System for General English Proficiency of College Students (Dong Seo) | Using Semantic Embeddings to Examine Content Overlap Between the IDCP-2 and PID-5 (Pedro Godoy dos Santos) | |
| Interactions Between Termination Criteria and Ability Estimators in Computerized Adaptive Testing (Xinyu Liu) | Growth Norms for Computer Adaptive Tests: Evaluating Three Conditional Growth Percentile Methods (Luciana Cancado) | AI-Based Item Parameter Estimation as an Alternative to Pilot Testing in Computerized Adaptive Testing: A Simulation Study (Zeus Bellido) | |
| One of these things is not like the other: Setting Thresholds for Equivalence of Item Response Theory Parameter Estimation (Advancing Adaptive Measurement) (Robert Chapman) | Enabling Computerized Adaptive Testing in large-scale educational assessments under limited connectivity conditions (Gustavo Silva) | Adaptive Testing Meets Contracting: Rethinking Success in the Age of AI (Xiangen Hu) | |
| Tracing the Roots and Branching into the Future: The Academic Legacy of Dr. David J. Weiss (Advancing Adaptive Measurement) (Robert Chapman) | Development of a Computerised Adaptive Test for Assessing Upper-Basic School Students? Mathematics-Ability in Nigeria: A Post-hoc Simulation Testing (ISAAC OLAWALE IFINJU) | AI-Assisted Cognitive Pretesting for Personality Assessment Items: Comparing Profile-Conditioned Simulated Responses With Observed Human Data (Diego Forteza) | |
| No session | Designing multi-stage CAT assessments of academic English language proficiency to give examinees the greatest opportunity to demonstrate that they meet performance level thresholds. (Stephen Walker) | Use of AI to Assess Metaphor Generation: Convergent And External Validity (Beatriz Chrispim) | |
| 9:40 - 10:00 | COFFEE BREAK | ||
| 10:00 - 11:00 |
Keynote 3: Hua-hua Chang Positioning Computerized Adaptive Testing as a Foundation for Personalized Learning |
||
| 11:00 - 12:00 |
Keynote 4: Maomi Ueno High-Stakes Assessments with Process-Integrated IRT and AI: From Empirical Evidence to Computational Design Strategies for Computerized Adaptive Testing |
||
| 12:00 - 14:00 | LUNCH | ||
| 14:00 - 15:40 | Individual presentations: Other AI Topics in Psychometrics | Symposia: New IRT models to modeling different responses | Symposia: Frontiers of Bayesian Adaptive Testing |
| Psychometrics Meets LLM: Diagnostic Assessment Copilot Design and Deployment (Chun Wang) | Revisiting the Skew-Normal IRT family: Bayesian estimation and application (Cristian Bayes) | Prior Specification for Ability Estimation in Bayesian Adaptive Testing (Jonathan Templin) | |
| Automated Evaluation of Teacher Portfolios Using Large Language Models and Structured Educational Evidence Analysis (Gustavo Silva) | A robust GNBk count item response model (Luis Valdivieso) | Spectral Efficiency in Multidimensional Bayesian Adaptive Testing: Comparing Fully Bayesian, Sequential Two-Level, and Globally Optimized Two-Level Frameworks (Seung Choi) | |
| Methodology for the Empirical Generation and Validation of Performance Level Descriptors Using Generative Artificial Intelligence (Daniela Jimenez) | Bayesian Estimation for a class of ability-based guessing IRT models (Alex de la Cruz Huayanay) | Bayesian Adaptive Testing for the Answer-Until-Correct-Response Format (Andreas Frey) | |
| Evaluating Test Fairness Using Machine Learning Procedures (Hanan AlGhamdi) | No session | Bayesian Optimal Large-Scale CAT Using Constraint-Governed Augmentation (Richard Patz) | |
| 15:40 - 16:10 | COFFEE BREAK | ||
| 16:30 - 21:30 | Departure for the Cathedral and Conference Dinner | ||
Wednesday, 30 September
- Morning: Plenary and Paper Sessions
- Afternoon: Closing Plenary and Paper Sessions
- Evening: Tour in São Paulo*
| ROOM 1 | ROOM 2 | ROOM 3 | |
|---|---|---|---|
| 8:00 - 9:00 |
Early Career Award: Dr Hyeon-Ah Kang Beyond Information: Balancing Measurement Precision, Testing Efficiency, and Item Security in Technology-Enhanced Adaptive Testing |
||
| 9:15 - 10:30 | Individual presentations: Cognitive diagnostic and IRT models (including multidimensional) | Symposia: Continuous Item Banking for a High-Stakes Certification Without Pretesting: An Operational Decomposition | Individual Presentations: Item generation, item banking, form assembly, and others |
| A Cognitive Diagnosis Model for Latent Classification of Bounded Continuous Variables (Eduardo Schneider Bueno de Oliveira) | Interactive Branching-Scenario Items in High-Stakes Financial Certification: Bundle Scoring Without Local Dependence (Andrea Burgos de Azevedo Mangabeira) | Automatic Pre-Testing of Mathematics Assessment Items: Predicting Item Response Theory (IRT) Difficulty Parameter with Machine Learning Methods (Patrick Canto de Carvalho) | |
| Black-Box Variational Inference with REINFORCE for the DINA/DINO Model (Flavio Barros) | Closing the Loop on Item Demand: A Decomposable Formalism for Continuous Item Banking (Marco Pepe) | Accelerating Item Piloting with NLP-Based Features and Bayesian CAT (Steven Nydick) | |
| Testing Differential Item Functioning in the Presence of Guessing (Hanan AlGhamdi) | Continuous Calibration of an Operational Item Bank Without Pretesting (Erica Ruiz) | LLM and Rule-Based Automatic Item Generation: A Two-Study Investigation in Large-Scale Literacy Assessment (Lucas Larcher) | |
| 10:30 - 10:45 | COFFEE BREAK | ||
| 10:45 - 12:00 | Symposia: The PROMIS, NIH Toolbox, and Mobile Toolbox Measurement Systems: Architecture, Psychometric Foundation, and the Challenges of Adaptive Health Measurement | Individual Presentations: Applications of CAT | Individual Presentations: Applications of CAT |
| Computerized Adaptive Testing (CAT) for Patient-Reported Outcomes Measurement Information System (PROMIS): Evaluation of Stopping Rules for the Pediatric Item Banks (Jiwon Kim) | AN IRT-BASED ADAPTIVE MEASURE OF CONSCIENTIOUSNESS IN THE BIG FIVE FRAMEWORK (Gabriela S. Lozzia) | Review and Requalification of a Workforce Competency Assessment Tool for Technical Education (Paloma de Lima Santos) | |
| NIH Toolbox Calibration and Scaling for a multi-stage episodic memory test across the lifespan (Emily Ho) | Adaptive Test of Vocational Interests (TIPA): Development and psychometric properties of a computerized assessment measure (Gustavo Martins) | Design and Implementation of Practice-Oriented, Large-Scale Certification Exams: The Case of ANBIMA's Distribution Certifications (Andrea Burgos de Azevedo Mangabeira) | |
| Computer Adaptive Testing on Mobile Devices: Assessments of Verbal Ability for Remote Administration in Mobile Toolbox (Y. Catherine Han) | Using Item Response Theory to Develop a Parsimonious Measure of Psychological Need Satisfaction at Work (Alexsandro Luiz de Andrade) | Enhancing Validity and Data Quality in Skills and Competency Assessments by Identifying Individuals Who Compromise the Measurement Process (Hanan AlGhamdi) | |
| 12:00 - 14:00 | LUNCH | ||
| 14:00 - 15:40 | Symposia: EduCAT | Individual Presentations: MST, Item generation, item banking, form assembly, and others | Individual Presentations: Statistical methodology |
| EduCAT | Form Assembly: An AI Agent Workflow for Building Parallel Spring Forms for a Diagnostic Classification Assessment (Ann Hu) | Beyond Correct and Incorrect: A Nested Logit Computerized Adaptive Test Using Distractor Information to Assess Reading Proficiency (Thiago Costa) | |
| A Stochastic Constrained Test Assembly Method Applied to Multistage Adaptive Testing (Alina von Davier) | Computerized Adaptive Testing Using Nonparametric Item Response Theory Models (Cecilia Marconi) | ||
| Designing multi-stage CAT assessments of academic English language proficiency to give examinees the greatest opportunity to demonstrate that they meet performance level thresholds (Stephen Walker) | Addressing Misclassification Bias in Computerized Adaptive Testing with Latent Performance Standards (Bartosz Kondratek) | ||
| parATA: an R-package for Automated Test Assembly (Angela Verschoor) | Testlet-Based Termination in Relaxed CAT: Reducing Assessment Length and Testlet Exposure (Luciana Cancado) | ||
| No session | Computer Adaptive Testing with small item bank: Improved precision and reduced opportunities for cheating (Rense Lange) | ||
| 15:40 - 16:10 | COFFEE BREAK | ||
| 16:10 - 17:10 |
Keynote 5: Ricardo Primi Navigating the Psychometric AI Revolution: Innovations in Scoring and Validity |
||
| 17:10 - 17:30 | Ending Ceremony | ||
Thursday-Friday, 1-2 October
- Organized tour in Rio de Janeiro*
*Additional Options
Pre-conference Workshops
Introduction to IRT and CAT
Nathan Thompson (ASC)
AI Applications in Adaptive Testing
Duanli Yan (ETS)
Alina A. von Davier (Duoling)
The Shadow-Test Approach to Adaptive Testing
Seung W. Choi (UT-Austin)
Richard J. Patz (University of California-Berkeley)
Wim J. van der Linden (University of Twente)
Introduction to IRT and CAT
Nathan Thompson (ASC)
New to CAT? This workshop is for you. We will start with an introduction to the CAT algorithm, including item selection, exposure constraints, scoring, and termination rules. We will discuss how this algorithm can be adapted to various approaches, including multistage testing, and how to utilize simulations to design and validate your CAT. Workshop assumes you are familiar with item response theory, and ready to apply it in adaptive testing.
AI Applications in Adaptive Testing
Duanli Yan (ETS)
Alina A. von Davier (Duoling)
In the era of artificial intelligence (AI), the field of educational testing faces significant challenges, particularly in test development and scoring in adaptive assessment, two key innovations include automated item generation (AIG) and automated scoring (AS). Only recently that generative AI has facilitated the development of complex test items on a large scale. We introduce “the item factory”, for managing large-scale test development including automation of item generation, quality review, quality assurance, and crowdsourcing techniques in adaptive testing. We present an overview of the latest natural language processing (NLP) techniques and large language models for AIG, alongside psychometric principles and practices for test development. We discuss the application of engineering principles in designing efficient item production processes (Luecht, 2008; Dede et al, 2018; von Davier, 2017). As AS becomes an integral part of the assessment landscape due to their advantages in reporting time, cost, objectivity, consistency, transparency, and feedback. We aim to demystify AS and provide a comprehensive understanding of its workings. We offer an overview of the design, development, evaluation, and quality control of automated scoring systems, along with practical advice and considerations for practitioners on the applications of these systems into formative and summative assessments (Yan, Rupp, & Foltz, 2020). We will share the most recent AI applications in educational learning and assessment based on our upcoming volume AI for Measurement in Educational Learning and Assessment (von Davier and Yan, 2026).
The Shadow-Test Approach to Adaptive Testing
Seung W. Choi (UT-Austin)
Richard J. Patz (University of California-Berkeley)
Wim J. van der Linden (University of Twente)
It is tempting to think of adaptive testing as a more complicated version of the problem of automated assembly of fixed test forms. But thanks to a simple twist of the latter, the problem is easier to solve, always produces maximum information about the ability of each of the test takers, meets any blueprint in force for the test, and easily generalizes to testing formats with varying degrees of adaptation such as item-level adaptive testing, linear-on-the-fly testing, standard multistage adaptive testing, and multistage testing with adaptive routing tests.
The course has three different parts. In the first part, we explain the ideas underlying the shadow-test approach, discuss a few practical aspects of its implementation, and show some of its generalizations to different test formats. The next two parts are to introduce two software packages available for the implementation of the shadow-test approach to adaptive testing, offering the participants hands-on experience with the R package TestDesign and Optimal CAT, a currently freely downloadable microservice available for easy integration with common test delivery systems.
Participants are expected to bring their own laptops to the course. Handouts will be sent to the participants prior to the conference to prepare for the course and review its content afterwards.
Full Program
The full program will be published as a PDF in August.
