Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- James Madison University (14)
- University of Nebraska - Lincoln (7)
- Pace University (6)
- Australian Council for Educational Research (ACER) (2)
- Wayne State University (2)
-
- American University in Cairo (1)
- Antioch University (1)
- Association of Arab Universities (1)
- City University of New York (CUNY) (1)
- Columbus State University (1)
- Florida Institute of Technology (1)
- Illinois State University (1)
- National Council on Measurement in Education (1)
- Northern Illinois University (1)
- United Arab Emirates University (1)
- University of Dar es Salaam (1)
- University of Denver (1)
- University of South Florida (1)
- Walden University (1)
- Western Michigan University (1)
- Keyword
-
- Assessment (4)
- Psychometrics (4)
- 3PL (2)
- DIF (2)
- IRT (2)
-
- MIRT (2)
- Measurement (2)
- Missing data (2)
- Multidimensional item response theory (2)
- Multilevel (2)
- Multiple imputation (2)
- Reliability (2)
- Simulation study (2)
- Validity (2)
- AAVE (1)
- AI Literacy (1)
- Ability distribution (1)
- Academic Performance (1)
- Academic Progress Indicator (API) (1)
- Accountability systems (1)
- Adolescents (1)
- Application-based training (1)
- Applied statistics, Bonferroni, dependent t-test, Maximum test, mixed normal distribution, Wilcoxon sign rank test (1)
- Attitude-achievement paradox (1)
- Attitudes (1)
- Auditory attention (1)
- Bayes Factor (1)
- Bayesian inference (1)
- Bias (1)
- Bifactor (1)
- Publication Year
- Publication
-
- Perspectives on Early Childhood Psychology and Education (6)
- Buros Center: Professional Staff Publications (4)
- Masters Theses, 2010-2019 (4)
- Department of Graduate Psychology - Faculty Scholarship (3)
- Dissertations, 2014-2019 (3)
-
- Theses and Dissertations (3)
- College of Education and Human Sciences: Dissertations, Theses, and Student Research (2)
- Dissertations, 2020-current (2)
- International Conference on Assessment and Learning (ICAL) (2)
- Masters Theses, 2020-current (2)
- Wayne State University Dissertations (2)
- An-Najah University Journal for Research - B (Humanities) (1)
- Antioch University Dissertations & Theses (1)
- Chinese/English Journal of Educational Measurement and Evaluation | 教育测量与评估双语期刊 (1)
- Department of Educational Psychology: Faculty Publications (1)
- Dissertations (1)
- Dissertations, Theses, and Capstone Projects (1)
- Electronic Theses and Dissertations (1)
- International Journal for Research in Education (1)
- Journal of Humanities and Social Sciences (1)
- Perspectives In Learning (1)
- Reports, Analyses, and other Publications (1)
- USF Tampa Graduate Theses and Dissertations (1)
- Walden Dissertations and Doctoral Studies (1)
- Publication Type
Articles 31 - 46 of 46
Full-Text Articles in Quantitative Psychology
Applying Conditional Distributions To Individuals: Using Latent Variable Models, Feng Ji
Applying Conditional Distributions To Individuals: Using Latent Variable Models, Feng Ji
Theses and Dissertations
This study proposes a new method to interpret individual results of psychological test batteries. The Mahalanobis distance is a commonly-used measure of how unusual an individual’s profile of scores is compared to a population of score profiles. In models in which there is a set of predictors and a set of dependent variables (e.g., cognitive abilities predicting academic abilities), it is useful to distinguish between a profile of dependent scores that is unusual because its profile of predictor scores is unusual and a profile of dependent scores that is unusual even after controlling for the predictors. The conditional Mahalanobis distance …
Posterior Predictive Model Checking Of Local Misfit For Bayesian Confirmatory Factor Analysis, Chi Hang Au
Posterior Predictive Model Checking Of Local Misfit For Bayesian Confirmatory Factor Analysis, Chi Hang Au
Masters Theses, 2010-2019
Posterior predictive model checks (PPMC) are one Bayesian model-data fit approach. Thus far, PPMC for Confirmatory Factor Analytic applications focused primarily on global fit evaluation, ignoring the nuanced information in local misfit diagnostics. This study developed a PPMC approach for local misfit and applied it to a test-taking motivation scale. If the PPMC approach is effective, fit conclusions derived from the PPMC approach should be congruent with the fit conclusions derived from the Frequentist approach. Number of item-pairs flagged as misfitting and number of disagreements were computed to evaluate congruence. Congruence is achieved if the number of item-pairs flagged as …
The Trouble With Test Banks, Harvey Richman, Molly Hrezo
The Trouble With Test Banks, Harvey Richman, Molly Hrezo
Perspectives In Learning
We compared the psychometrics of quiz questions randomly selected from a test bank with the psychometrics of quiz questions the instructor had selected from the bank for quality and modified (if necessary). On multiple psychometric indices, the instructor selected/modified questions were superior to questions randomly selected from the test bank. Most notably, when compared with instructor written/modified questions, randomly selected bank questions were nearly 6.5 times more likely to contain a distractor that drew more responses than the correct answer. Details and implications are discussed.
Retrospective Versus Prospective Measurement Of Examinee Motivation In Low-Stakes Testing Contexts: A Moderated Mediation Model, Aaron J. Myers
Retrospective Versus Prospective Measurement Of Examinee Motivation In Low-Stakes Testing Contexts: A Moderated Mediation Model, Aaron J. Myers
Masters Theses, 2010-2019
Expectancy-value theory applied to examinee motivation suggests examinees’ perceived value of a test indirectly affects test performance via examinee effort. This empirically supported indirect effect, however, is often modeled using importance and effort scores measured after test completion, which does not align with their theoretically specified temporal order. Retrospectively measured importance and effort scores may be influenced by examinees’ test performance, impacting the estimate of the indirect effect. To investigate the effect of timing of measurement, first-year college students were randomly assigned to one of three conditions where (1) importance and effort were measured retrospectively; (2) importance was measured prospectively; …
Student Learning Gains In Higher Education: A Longitudinal Analysis With Faculty Discussion, Catherine E. Mathers
Student Learning Gains In Higher Education: A Longitudinal Analysis With Faculty Discussion, Catherine E. Mathers
Masters Theses, 2010-2019
Student learning is the primary desired outcome of a college education. To understand how educational programming and curricula affect students, colleges and universities must collect evidence of student learning gain. In this study, a longitudinal design was employed to investigate how a math and science general education curriculum impacted college students’ quantitative and scientific reasoning. Quantitative and scientific reasoning gain scores were computed and predicted from personal (i.e., prior knowledge, gender) and curriculum (i.e., number of completed courses in the domain) characteristics to uncover what factors relate to learning gain. Collapsing across personal and curriculum variables, gain scores were moderate …
You Only Live Up To The Standards You Set: An Evaluation Of Different Approaches To Standard Setting, Scott N. Strickman
You Only Live Up To The Standards You Set: An Evaluation Of Different Approaches To Standard Setting, Scott N. Strickman
Dissertations, 2014-2019
Interpretation of performance in reference to a standard can provide nuanced, finely-tuned information regarding examinee abilities beyond that of just a total score. However, there is a multitude of ways to set performance standards yet little guidance regarding which method operates best and under what circumstances. Traditional methods are the most common approach adopted in practice and heavily involve subject matter experts (SMEs). Two other approaches have been suggested in the literature as alternative ways to set performance standards, although they have yet to be implemented in practice. Data-driven approaches do not involve SMEs but rather rely solely upon statistical …
Strategies And Resources To Enhance Test Evaluation And Selection, Janet F. Carlson, Nancy Anderson
Strategies And Resources To Enhance Test Evaluation And Selection, Janet F. Carlson, Nancy Anderson
Buros Center: Professional Staff Publications
Testing serves an important function for SLPs in offering an evidence base that is useful in screening, diagnosing, monitoring progress, and documenting outcomes. Tests are used to measure diverse constructs such as communication, literacy, oral and written language, receptive and expressive vocabulary, articulation, phonological awareness and processing, and auditory perception and processing. In addition, specific impairments may require specialized measures to evaluate conditions such as stuttering and orthographic competence.
When using tests to diagnose language impairments, Betz, Eickhoff, and Sullivan (2013) suggest that SLPs consider carefully a test’s psychometric properties, particularly because of the “increasing emphasis on evidence-based practice, specifically, …
The Effects Of A Planned Missingness Design On Examinee Motivation And Psychometric Quality, Matthew S. Swain
The Effects Of A Planned Missingness Design On Examinee Motivation And Psychometric Quality, Matthew S. Swain
Dissertations, 2014-2019
Assessment practitioners in higher education face increasing demands to collect assessment and accountability data to make important inferences about student learning and institutional quality. The validity of these high-stakes decisions is jeopardized, particularly in low-stakes testing contexts, when examinees do not expend sufficient motivation to perform well on the test. This study introduced planned missingness as a potential solution. In planned missingness designs, data on all items are collected but each examinee only completes a subset of items, thus increasing data collection efficiency, reducing examinee burden, and potentially increasing data quality. The current scientific reasoning test served as the Long …
Examining The Performance Of The Metropolis-Hastings Robbins-Monro Algorithm In The Estimation Of Multilevel Multidimensional Irt Models, Bozhidar M. Bashkov
Examining The Performance Of The Metropolis-Hastings Robbins-Monro Algorithm In The Estimation Of Multilevel Multidimensional Irt Models, Bozhidar M. Bashkov
Dissertations, 2014-2019
The purpose of this study was to review the challenges that exist in the estimation of complex (multidimensional) models applied to complex (multilevel) data and to examine the performance of the recently developed Metropolis-Hastings Robbins-Monro (MH-RM) algorithm (Cai, 2010a, 2010b), designed to overcome these challenges and implemented in both commercial and open-source software programs. Unlike other methods, which either rely on high-dimensional numerical integration or approximation of the entire multidimensional response surface, MH-RM makes use of Fisher’s Identity to employ stochastic imputation (i.e., data augmentation) via the Metropolis-Hastings sampler and then apply the stochastic approximation method of Robbins and Monro …
Validation Of The Item-Attribute Matrix In Timss-Mathematics Using Multiple Regression And The Lsdm, Lin Ma
Validation Of The Item-Attribute Matrix In Timss-Mathematics Using Multiple Regression And The Lsdm, Lin Ma
Electronic Theses and Dissertations
For many cognitive diagnostic models, the item-attribute matrix (or Q-matrix) is an essential component which displays the relationship between items and their latent attributes or skills in knowledge and cognitive processes. However, it is a challenge to develop an effective Q-matrix.The purposes of this study were (1) to validate of the item-attribute matrix using two levels of attributes (Level 1 attributes and Level 2 sub-attributes), and (2) through retrofitting the diagnostic models to the mathematics test of the Trends in International Mathematics and Science Study (TIMSS), to evaluate the construct validity of TIMSS mathematics assessment by comparing the results of …
The Reliability And Validity Of The Thin Slice Technique: Observational Research On Video Recorded Medical Interactions, Tanina Suzanne Foster
The Reliability And Validity Of The Thin Slice Technique: Observational Research On Video Recorded Medical Interactions, Tanina Suzanne Foster
Wayne State University Dissertations
The Reliability and Validity of the Thin Slice Technique: Observational Research on Video Recorded Medical Interactions
Introduction: Observational research using the thin slice technique has been routinely incorporated in observational research methods, however there is limited evidence supporting use of this technique compared to full interaction coding. The purpose of this study was to determine if this technique could be reliability coded, if ratings are consistent between the first, second and third slice, and if they are indeed representative of full interactions.
Methods: Three 30-second thin slices were sampled from the beginning, middle and end of a full-length video-recorded …
Teacher Support Mediates Concurrent And Longitudinal Associations Between Temperament And Mild Depressive Symptoms In Sixth Grade, Kathleen Moritz Rudasill, Patrick Pössel, Stephanie Winkeljohn Black, Kate Niehaus
Teacher Support Mediates Concurrent And Longitudinal Associations Between Temperament And Mild Depressive Symptoms In Sixth Grade, Kathleen Moritz Rudasill, Patrick Pössel, Stephanie Winkeljohn Black, Kate Niehaus
Department of Educational Psychology: Faculty Publications
The combination of changes occurring at the transition to middle school may be a catalyst for the onset of depressive symptoms, yet teacher support at this transition is protective. Research points to certain temperamental traits as risk factors for developing depressive symptoms. This study examines student reports of teacher support and teacher reports of student–teacher relationship (STR) quality as mediators of associations between child temperament (i.e. negative emotionality at age 4½ : and emotional reactivity in elementary grades) and depressive symptoms in sixth grade. Results indicate (a) negative emotionality predicted emotional reactivity and depressive symptoms; (b) emotional reactivity predicted depressive …
An Analysis Of Factor Extraction Strategies: A Comparison Of The Relative Strengths Of Principal Axis, Ordinary Least Squares, And Maximum Likelihood In Research Contexts That Include Both Categorical And Continuous Variables, Kevin Barry Coughlin
USF Tampa Graduate Theses and Dissertations
This study is intended to provide researchers with empirically derived guidelines for conducting factor analytic studies in research contexts that include dichotomous and continuous levels of measurement. This study is based on the hypotheses that ordinary least squares (OLS) factor analysis will yield more accurate parameter estimates than maximum likelihood (ML) and principal axis factor anlaysis (PAF); the level of improvement in estimates will be related to the proportion of observed variables that are dichotomized and the strength of communalities within the data sets.
To achieve this study's objective, maximum likelihood, ordinary least squares, and principal axis factor extraction models …
A Preliminary Investigation Of The Validity Of Time-Based Measures Of Sustained Attention For Children, Michael R. Kulfan
A Preliminary Investigation Of The Validity Of Time-Based Measures Of Sustained Attention For Children, Michael R. Kulfan
Antioch University Dissertations & Theses
This study is a preliminary investigation of the validity of using time-based measures to quantify sustained attention in children ages 6-12. Problems with sustained attention negatively affect childhood learning and development. The prevalence of disorders known to impact sustained attention performance continue to rise in the United States. Currently, commercially available, objective measures of sustained attention use normative comparisons that provide limited information about the effect such problems have on child performance in natural settings. We reviewed test data from 290 charts of children ages 6-12 referred for neuropsychological evaluation. The Test of Everyday Attention for Children (TEA-Ch) is an …
The Dependent Samples T And Wilcoxon Sign Rank Maximum Test, Saverpierre Maggio
The Dependent Samples T And Wilcoxon Sign Rank Maximum Test, Saverpierre Maggio
Wayne State University Dissertations
A maximum test using the parametric dependent samples t-test and the non-parametric Wilcoxon sign rank test was created using a FORTRAN program and various subroutines of the International Mathematical and Statistical Libraries (IMSL, 1980). Two tailed critical values were derived from a mixed normal distribution. Critical values obtained were at the 0.05, 0.025, 0.01 and 0.005 alpha levels via sample sizes (n) 8 through 30, 45, 60, 90 and 120. Critical values were compared to values obtained through the application of the Bonferroni correction method. It was concluded that the Bonferroni is an unnecessary method. Findings of the study are …
Improving Irt Parameter Estimates With Small Sample Sizes: Evaluating The Efficacy Of A New Data Augmentation Technique, Brett P. Foley
Improving Irt Parameter Estimates With Small Sample Sizes: Evaluating The Efficacy Of A New Data Augmentation Technique, Brett P. Foley
College of Education and Human Sciences: Dissertations, Theses, and Student Research
The 3PL model is a flexible and widely used tool in assessment. However, it suffers from limitations due to its need for large sample sizes. This study introduces and evaluates the efficacy of a new sample size augmentation technique called Duplicate, Erase, and Replace (DupER) Augmentation through a simulation study. Data are augmented using several variations of DupER Augmentation (based on different imputation methodologies, deletion rates, and duplication rates), analyzed in BILOG-MG 3, and results are compared to those obtained from analyzing the raw data. Additional manipulated variables include test length and sample size. Estimates are compared using seven different …