ERIC - Search Results

Publication Date

In 2025	6
Since 2024	14
Since 2021 (last 5 years)	55
Since 2016 (last 10 years)	140
Since 2006 (last 20 years)	259

Descriptor

Scoring	259
Test Reliability	259
Test Validity	176
Test Construction	87
Test Items	69
Testing	58
Psychometrics	51
Item Response Theory	49
Foreign Countries	45
Scores	45
Language Tests	39
Computer Assisted Testing	36
Elementary School Students	31
Mathematics Tests	31
Correlation	30
Test Bias	29
Interrater Reliability	28
Language Arts	26
Comparative Analysis	25
Item Analysis	23
Children	21
Error of Measurement	21
Test Interpretation	21
Factor Analysis	20
Grade 3	20
More ▼

Publication Type

Journal Articles	185
Reports - Research	133
Reports - Evaluative	78
Reports - Descriptive	26
Numerical/Quantitative Data	23
Tests/Questionnaires	15
Guides - Non-Classroom	9
Books	5
Dissertations/Theses -…	5
Speeches/Meeting Papers	5
Guides - General	4
Collected Works - General	2
Guides - Classroom - Teacher	2
Information Analyses	2
Opinion Papers	2
More ▼

Education Level

Elementary Education	52
Secondary Education	47
Higher Education	36
Postsecondary Education	31
Middle Schools	29
Early Childhood Education	28
Junior High Schools	26
Elementary Secondary Education	23
Primary Education	22
Grade 3	20
High Schools	20
Grade 4	19
Grade 5	19
Grade 8	19
Intermediate Grades	19
Grade 7	18
Grade 6	17
Kindergarten	9
Grade 1	7
Grade 2	6
Preschool Education	6
Grade 9	5
Grade 11	4
Grade 10	3
Grade 12	2
More ▼

Audience

Administrators	6
Practitioners	3
Teachers	3
Policymakers	2

Location

New York	14
Turkey	8
Nebraska	7
Florida	6
California	4
Canada	4
Netherlands	3
New Mexico	3
Texas	3
United Kingdom	3
United States	3
Europe	2
Germany	2
Taiwan	2
United Kingdom (England)	2
Alabama	1
Australia	1
California (Los Angeles)	1
Chile	1
District of Columbia	1
Estonia	1
Finland	1
Idaho	1
India	1
Indonesia	1
More ▼

Laws, Policies, & Programs

Individuals with Disabilities…	5
No Child Left Behind Act 2001	2
Individuals with Disabilities…	1
Individuals with Disabilities…	1

What Works Clearinghouse Rating

Meets WWC Standards without Reservations	1
Meets WWC Standards with or without Reservations	1

Showing 1 to 15 of 259 results Save | Export

A Study on Psychometric Properties of Creativity Indices

Peer reviewed

Direct link

M. Arda Atakaya; Ugur Sak; M. Bahadir Ayas – Creativity Research Journal, 2024

Scoring in creativity research has been a central problem since creativity became an important issue in psychology and education in the 1950s. The current study examined the psychometric properties of 27 creativity indices derived from summed and averaged scores using 15 scoring methods. Participants included 2802 middle-school students. Data…

Descriptors: Psychometrics, Creativity, Creativity Tests, Scoring

The Sensitivity of Value-Added Estimates to Test Scoring Decisions. EdWorkingPaper No. 25-1226

Download full text

Joshua B. Gilbert; James G. Soland; Benjamin W. Domingue – Annenberg Institute for School Reform at Brown University, 2025

Value-Added Models (VAMs) are both common and controversial in education policy and accountability research. While the sensitivity of VAMs to model specification and covariate selection is well documented, the extent to which test scoring methods (e.g., mean scores vs. IRT-based scores) may affect VA estimates is less studied. We examine the…

Descriptors: Value Added Models, Tests, Testing, Scoring

Linking Errors Introduced by Rapid Guessing Responses When Employing Multigroup Concurrent IRT Scaling

Direct link

Jiayi Deng – ProQuest LLC, 2024

Test score comparability in international large-scale assessments (LSA) is of utmost importance in measuring the effectiveness of education systems and understanding the impact of education on economic growth. To effectively compare test scores on an international scale, score linking is widely used to convert raw scores from different linguistic…

Descriptors: Item Response Theory, Scoring Rubrics, Scoring, Error of Measurement

Is Effort Moderated Scoring Robust to Multidimensional Rapid Guessing?

Peer reviewed

Direct link

Joseph A. Rios; Jiayi Deng – Educational and Psychological Measurement, 2025

To mitigate the potential damaging consequences of rapid guessing (RG), a form of noneffortful responding, researchers have proposed a number of scoring approaches. The present simulation study examines the robustness of the most popular of these approaches, the unidimensional effort-moderated (EM) scoring procedure, to multidimensional RG (i.e.,…

Descriptors: Scoring, Guessing (Tests), Reaction Time, Item Response Theory

Do Scoring Techniques and Number of Choices Affect the Reliability of Multiple-Choice Tests in Elementary Schools?

Peer reviewed
PDF on ERIC

Download full text

Herwin, Herwin; Pristiwaluyo, Triyanto; Ruslan, Ruslan; Dahalan, Shakila Che – Cypriot Journal of Educational Sciences, 2022

The application of multiple-choice tests often does not consider the scoring technique and the number of choices. The study aims at describing the effect of the scoring technique and numerous options towards the reliability of multiple-choice objective tests on social subjects in elementary school. The study is quantitative research with…

Descriptors: Scoring, Multiple Choice Tests, Test Reliability, Elementary School Students

Evaluating the Consistency and Reliability of Attribution Methods in Automated Short Answer Grading (ASAG) Systems: Toward an Explainable Scoring System

Peer reviewed

Direct link

Wallace N. Pinto Jr.; Jinnie Shin – Journal of Educational Measurement, 2025

In recent years, the application of explainability techniques to automated essay scoring and automated short-answer grading (ASAG) models, particularly those based on transformer architectures, has gained significant attention. However, the reliability and consistency of these techniques remain underexplored. This study systematically investigates…

Descriptors: Automation, Grading, Computer Assisted Testing, Scoring

Computational Concepts and Their Assessment in Preschool Students: An Empirical Study

Peer reviewed

Direct link

Marcos Jiménez; María Zapata-Cáceres; Marcos Román-González; Gregorio Robles; Jesús Moreno-León; Estefanía Martín-Barroso – Journal of Science Education and Technology, 2024

Computational thinking (CT) is a multidimensional term that encompasses a wide variety of problem-solving skills related to the field of computer science. Unfortunately, standardized, valid, and reliable methods to assess CT skills in preschool children are lacking, compromising the reliability of the results reported in CT interventions. To…

Descriptors: Computation, Thinking Skills, Student Evaluation, Preschool Children

Selecting Technically Adequate Tests

Peer reviewed

Direct link

Susan K. Johnsen – Gifted Child Today, 2024

The author provides a checklist for educators who are selecting technically adequate tests for identifying and referring students for gifted education services and programs. The checklist includes questions related to how the test was normed, reliability and validity studies as well as questions related to types of scores, administration, and…

Descriptors: Test Selection, Academically Gifted, Gifted Education, Test Validity

Comparison of the Results of the Generalizability Theory with the Inter-Rater Agreement Coefficients

Peer reviewed
PDF on ERIC

Download full text

Eser, Mehmet Taha; Aksu, Gökhan – International Journal of Curriculum and Instruction, 2022

The agreement between raters is examined within the scope of the concept of "inter-rater reliability". Although there are clear definitions of the concepts of agreement between raters and reliability between raters, there is no clear information about the conditions under which agreement and reliability level methods are appropriate to…

Descriptors: Generalizability Theory, Interrater Reliability, Evaluation Methods, Test Theory

Preservice Teachers' Knowledge of Math Modeling: Initial Scale Development and Validation

Peer reviewed

Direct link

Reuben S. Asempapa; Doris Lee – Discover Education, 2025

Across the world, standards and practices for preparing teachers of mathematics emphasize the importance of math modeling (MM) in developing students' mathematical thinking. The aim of this research study was to develop the Mathematical Modeling Knowledge Scale (MAMKS), capable of determining preservice teachers' (PSTs') knowledge of MM. The study…

Descriptors: Preservice Teachers, Preservice Teacher Education, Mathematics Education, Mathematics Curriculum

A Review of Test Use: The Test Anxiety Inventory

Peer reviewed
PDF on ERIC

Download full text

Alatli, Betül – International Journal of Curriculum and Instruction, 2022

This study was conducted to review the use of tests. For this purpose, 45 articles in which the Turkish form of the "Test Anxiety Inventory (TAI)," which is one of the tests frequently used in the field of education, was employed and that were published between 2000 and 2020 were examined in terms of factors that should be considered in…

Descriptors: Anxiety, Likert Scales, Test Anxiety, Test Reliability

A Novel Automated Essay Scoring Approach for Reliable Higher Educational Assessments

Peer reviewed

Direct link

Beseiso, Majdi; Alzubi, Omar A.; Rashaideh, Hasan – Journal of Computing in Higher Education, 2021

E-learning is gradually gaining prominence in higher education, with universities enlarging provision and more students getting enrolled. The effectiveness of automated essay scoring (AES) is thus holding a strong appeal to universities for managing an increasing learning interest and reducing costs associated with human raters. The growth in…

Descriptors: Automation, Scoring, Essays, Writing Tests

Assessing Handwriting in Preschool-Aged Children: Reliability and Internal Consistency of the "Just Write!" Tool

Peer reviewed

Direct link

Bolton, Tiffany; Stevenson, Brittney; Janes, William – Journal of Occupational Therapy, Schools & Early Intervention, 2023

Researchers utilized a cross-sectional secondary analysis of data within an ongoing non-randomized controlled trial study design to establish the reliability and internal consistency of a novel handwriting assessment for preschoolers, the Just Write! (JW), written by the authors. Seventy-eight children from an area preschool participated in the…

Descriptors: Handwriting, Writing Skills, Writing Evaluation, Preschool Children

Item Response Theory Modeling of the Verb Naming Test

Peer reviewed

Direct link

Fergadiotis, Gerasimos; Casilio, Marianne; Dickey, Michael Walsh; Steel, Stacey; Nicholson, Hannele; Fleegle, Mikala; Swiderski, Alexander; Hula, William D. – Journal of Speech, Language, and Hearing Research, 2023

Purpose: Item response theory (IRT) is a modern psychometric framework with several advantageous properties as compared with classical test theory. IRT has been successfully used to model performance on anomia tests in individuals with aphasia; however, all efforts to date have focused on noun production accuracy. The purpose of this study is to…

Descriptors: Item Response Theory, Psychometrics, Verbs, Naming

Initial Evidence Supporting Interpretations of Scores from the Enhanced ACT Test. ACT Research. Research Report. R2425

Download full text

Jeff Allen; Ty Cruce – ACT Education Corp., 2025

This report summarizes some of the evidence supporting interpretations of scores from the enhanced ACT, focusing on reliability, concurrent validity, predictive validity, and score comparability. The authors argue that the evidence presented in this report supports the interpretation of scores from the enhanced ACT as measures of high school…

Descriptors: College Entrance Examinations, Testing, Change, Scores

Previous Page | Next Page »

Pages: 1 | 2 | 3 | 4 | 5 | 6 | 7 | 8 | 9 | 10 | 11 | ... | 18

Journal of Psychoeducational…	41
New York State Education…	12
ETS Research Report Series	11
Canadian Journal of School…	8
Grantee Submission	8
Language Testing	7
Nebraska Department of…	6
Online Submission	5
ProQuest LLC	5
Educational Measurement:…	4
Partnership for Assessment of…	4
Assessment in Education:…	3
Educational and Psychological…	3
European Journal of…	3
Journal of Autism and…	3
Journal of Speech, Language,…	3
Language Assessment Quarterly	3
Measurement and Evaluation in…	3
Advances in Health Sciences…	2
Applied Measurement in…	2
Assessing Writing	2
Assessment for Effective…	2
Bill & Melinda Gates…	2
College Board	2
Early Education and…	2
More ▼

Schoen, Robert C.	7
McCrimmon, Adam W.	6
Yang, Xiaotong	4
Anderson, Daniel	3
Attali, Yigal	3
Bauduin, Charity	3
Paek, Insu	3
Allen, Abigail A.	2
Anna-Maria Fall	2
Bae, Yunhee	2
Balkin, Richard S.	2
Barbot, Baptiste	2
Beula M. Magimairaj	2
Climie, Emma A.	2
Dombrowski, Stefan C.	2
Espin, Christine A.	2
Fraccaro, Rebecca L.	2
Greg Roberts	2
Guo, Hongwen	2
Guthrie, Donald	2
Holling, Heinz	2
Jiayi Deng	2
Kane, Thomas J.	2
Kyllonen, Patrick	2
Lembke, Erica S.	2
More ▼

Wechsler Intelligence Scale…	7
ACT Assessment	5
Test of English as a Foreign…	5
Wechsler Individual…	4
Kaufman Test of Educational…	3
SAT (College Admission Test)	3
Wechsler Adult Intelligence…	3
Woodcock Johnson Tests of…	3
Beery Developmental Test of…	2
Clinical Evaluation of…	2
Graduate Record Examinations	2
Raven Progressive Matrices	2
Wechsler Preschool and…	2
ACT Interest Inventory	1
Advanced Placement…	1
Autism Diagnostic Observation…	1
Battelle Developmental…	1
Bayley Scales of Infant…	1
Behavior Assessment System…	1
Block Design Test	1
British Ability Scales	1
Childrens Depression Inventory	1
Cognitive Assessment System	1
Computer Attitude Scale	1
Conners Rating Scales	1
More ▼