ERIC - Search Results

Publication Date

In 2025	10
Since 2024	22

Publication Type

Journal Articles	20
Reports - Research	20
Tests/Questionnaires	2
Dissertations/Theses -…	1
Reports - Evaluative	1

Education Level

Higher Education	8
Postsecondary Education	8
Secondary Education	7
Middle Schools	3
Elementary Education	2
Elementary Secondary Education	2
High Schools	2
Intermediate Grades	2
Junior High Schools	2
Early Childhood Education	1
Grade 4	1
Grade 6	1
Grade 8	1
Preschool Education	1
More ▼

Audience

Location

Indonesia	3
Turkey	2
Bosnia and Herzegovina	1
Chile	1
Ecuador	1
Mexico	1
Panama	1
Pennsylvania (Pittsburgh)	1
Philippines	1
Spain	1
Thailand	1
Thailand (Bangkok)	1
United Kingdom	1
More ▼

Laws, Policies, & Programs

Assessments and Surveys

Test of English for…	1
Trends in International…	1
Watson Glaser Critical…	1

What Works Clearinghouse Rating

Showing 1 to 15 of 22 results Save | Export

Identifying Difficult Questions and Student Difficulties in a Spanish Version of a Programming Assessment Instrument (SCS1)

Peer reviewed

Direct link

Camilo Vieira; Andrea Vásquez; Federico Meza; Roxana Quintero-Manes; Pedro Godoy – ACM Transactions on Computing Education, 2024

Currently, there is little evidence about how non-English-speaking students learn computer programming. For example, there are few validated assessment instruments to measure the development of programming skills, especially for the Spanish-speaking population. Having valid assessment instruments is essential to identify the difficulties of the…

Descriptors: Programming, Spanish Speaking, Translation, Test Validity

A Chi-Square Statistic for Testing the Equality of Distracters' Plausibility in Multiple-Choice Test Items

Download full text

Sherwin E. Balbuena – Online Submission, 2024

This study introduces a new chi-square test statistic for testing the equality of response frequencies among distracters in multiple-choice tests. The formula uses the information from the number of correct answers and wrong answers, which becomes the basis of calculating the expected values of response frequencies per distracter. The method was…

Descriptors: Multiple Choice Tests, Statistics, Test Validity, Testing

Evaluating Methodological Enhancements to the Yes/No Angoff Standard-Setting Method in Language Proficiency Assessment

Peer reviewed

Direct link

Tia M. Fechter; Heeyeon Yoon – Language Testing, 2024

This study evaluated the efficacy of two proposed methods in an operational standard-setting study conducted for a high-stakes language proficiency test of the U.S. government. The goal was to seek low-cost modifications to the existing Yes/No Angoff method to increase the validity and reliability of the recommended cut scores using a convergent…

Descriptors: Standard Setting, Language Proficiency, Language Tests, Evaluation Methods

Validation of an Elicited Imitation Test as a Measure of Korean Language Proficiency

Peer reviewed

Direct link

Hojung Kim; Changkyung Song; Jiyoung Kim; Hyeyun Jeong; Jisoo Park – Language Testing in Asia, 2024

This study presents a modified version of the Korean Elicited Imitation (EI) test, designed to resemble natural spoken language, and validates its reliability as a measure of proficiency. The study assesses the correlation between average test scores and Test of Proficiency in Korean (TOPIK) levels, examining score distributions among beginner,…

Descriptors: Korean, Test Validity, Test Reliability, Imitation

Empirically Deriving Cut Scores in the Positive Behavioral Interventions and Supports (PBIS) Tiered Fidelity Inventory (TFI) through a Bookmarking Process

Peer reviewed

Direct link

Jerin Kim; Kent McIntosh – Journal of Positive Behavior Interventions, 2025

We aimed to identify empirically valid cut scores on the positive behavioral interventions and supports (PBIS) Tiered Fidelity Inventory (TFI) through an expert panel process known as bookmarking. The TFI is a measurement tool to evaluate the fidelity of implementation of PBIS. In the bookmark method, experts reviewed all TFI items and item scores…

Descriptors: Positive Behavior Supports, Cutting Scores, Fidelity, Program Evaluation

Improvised Progressive Model Based on Automatic Calibration of Difficulty Level: A Practical Solution of Competitive-Based Examination

Peer reviewed

Direct link

Aditya Shah; Ajay Devmane; Mehul Ranka; Prathamesh Churi – Education and Information Technologies, 2024

Online learning has grown due to the advancement of technology and flexibility. Online examinations measure students' knowledge and skills. Traditional question papers include inconsistent difficulty levels, arbitrary question allocations, and poor grading. The suggested model calibrates question paper difficulty based on student performance to…

Descriptors: Computer Assisted Testing, Difficulty Level, Grading, Test Construction

Argument-Based Validation of Chulalongkorn University Language Institute (CULI) Test: A Rasch-Based Evidence Investigation

Peer reviewed

Direct link

Apichat Khamboonruang – Language Testing in Asia, 2025

Chulalongkorn University Language Institute (CULI) test was developed as a local standardised test of English for professional and international communication. To ensure that the CULI test fulfils its intended purposes, this study employed Kane's argument-based validation and Rasch measurement approaches to construct the validity argument for the…

Descriptors: Universities, Second Language Learning, Second Language Instruction, Language Tests

Assessing Lower-Secondary School Students' Critical Thinking Skills in Photosynthesis: A Rasch Model Approach

Peer reviewed
PDF on ERIC

Download full text

Suwita Suwita; Sulistyo Saputro; Sajidan Sajidan; Sutarno Sutarno – Journal of Baltic Science Education, 2024

The current study uses the Rasch Model to measure lower-secondary school students' critical thinking skills on photosynthesis topics. Critical thinking skills are considered essential in science education, but few valid and practical measurement instruments remain. The current study fills the gap by adapting the instrument from the Watson-Glaser…

Descriptors: Secondary School Students, Critical Thinking, Thinking Skills, Botany

Developing and Validating a Biological System Thinking Test for Middle School Students

Peer reviewed

Direct link

Ruying Li; Gaofeng Li – International Journal of Science and Mathematics Education, 2025

Systems thinking (ST) is an essential competence for future life and biology learning. Appropriate assessment is critical for collecting sufficient information to develop ST in biology education. This research offers an ST framework based on a comprehensive understanding of biological systems, encompassing four skills across three complexity…

Descriptors: Test Construction, Test Validity, Science Tests, Cognitive Tests

Validity and Reliability Analysis of a Socioscientific Issues-Based Critical Thinking Self-Assessment Instrument Using the Rasch Model

Peer reviewed
PDF on ERIC

Download full text

Y. Yokhebed; Rexy Maulana Dwi Karmadi; Luvia Ranggi Nastiti – Journal of Biological Education Indonesia (Jurnal Pendidikan Biologi Indonesia), 2025

Although self-assessment in critical thinking is thought to help students recognise their strengths and weaknesses, the reliability and validity of the assessment tool is still questionable, so a more objective evaluation is needed. Objective of this investigation is to assess the self-assessment tools in evaluating students' critical thinking…

Descriptors: Self Evaluation (Individuals), Critical Thinking, Science and Society, Test Validity

Understanding Resilience in Programming: A Scale Adaptation and Analysis of Individual Differences

Peer reviewed

Direct link

Busra Ozmen Yagiz; Ecenaz Alemdag – Education and Information Technologies, 2025

Resilience is a critical personality trait that allows one to deal with difficulties, learn from failures, and maintain a positive attitude during task performance. However, it has not been understudied in a complex and challenging educational domain. The current research intends to address this gap by analyzing the specific characteristics of…

Descriptors: Foreign Countries, Undergraduate Students, Resilience (Psychology), Programming

Development and Validation of the Student Perception of Academic Challenge Scale

Direct link

Jenna M. T. Vest – ProQuest LLC, 2024

This study focuses on creating a reliable and valid instrument to measure high school students' perceptions of academic challenge. The research is divided into four phases: qualitative analysis, item development, exploratory factor analysis (EFA), and validation. Initial data from college students' retrospective views and high school students'…

Descriptors: Test Construction, Test Validity, Student Attitudes, Academic Achievement

Design, Development, and Evaluation of the Organic Chemistry Representational Competence Assessment (ORCA)

Peer reviewed

Direct link

Lyniesha Ward; Fridah Rotich; Jeffrey R. Raker; Regis Komperda; Sachin Nedungadi; Maia Popova – Chemistry Education Research and Practice, 2025

This paper describes the design and evaluation of the Organic chemistry Representational Competence Assessment (ORCA). Grounded in Kozma and Russell's representational competence framework, the ORCA measures the learner's ability to "interpret," "translate," and "use" six commonly used representations of molecular…

Descriptors: Organic Chemistry, Science Tests, Test Construction, Student Evaluation

The Knowledge of Autism Questionnaire-UK: Development and Initial Psychometric Evaluation

Peer reviewed

Direct link

Sophie Langhorne; Nora Uglik-Marucha; Charlotte Broadhurst; Elena Lieven; Amelia Pearson; Silia Vitoratou; Kathy Leadbitter – Journal of Autism and Developmental Disorders, 2025

Tools to measure autism knowledge are needed to assess levels of understanding within particular groups of people and to evaluate whether awareness-raising campaigns or interventions lead to improvements in understanding. Several such measures are in circulation, but, to our knowledge, there are no psychometrically-validated questionnaires that…

Descriptors: Foreign Countries, Autism Spectrum Disorders, Questionnaires, Psychometrics

Incorporating Evidence-Based Gamification and Machine Learning to Assess Preschool Executive Function: A Feasibility Study

Peer reviewed

Direct link

Cassondra M. Eng; Aria Tsegai-Moore; Anna V. Fisher – Grantee Submission, 2024

Computerized assessments and digital games have become more prevalent in childhood, necessitating a systematic investigation of the effects of gamified executive function assessments on performance and engagement. This study examined the feasibility of incorporating gamification and a machine learning algorithm that adapts task difficulty to…

Descriptors: Preschool Children, Preschool Curriculum, Preschool Education, Preschool Tests

Previous Page | Next Page »

Pages: 1 | 2

Education and Information…	2
Language Testing in Asia	2
ACM Transactions on Computing…	1
Chemistry Education Research…	1
Grantee Submission	1
International Electronic…	1
International Journal of…	1
Journal of Applied Research…	1
Journal of Autism and…	1
Journal of Baltic Science…	1
Journal of Biological…	1
Journal of Positive Behavior…	1
Journal of Social Studies…	1
Language Testing	1
Large-scale Assessments in…	1
Online Submission	1
Physical Review Physics…	1
ProQuest LLC	1
Science Insights Education…	1
rEFLections	1
More ▼

Difficulty Level	22
Test Validity	22
Test Items	16
Test Reliability	13
Foreign Countries	12
Test Construction	9
Science Tests	6
Multiple Choice Tests	5
Undergraduate Students	5
Factor Analysis	4
Item Response Theory	4
Psychometrics	4
Questionnaires	4
Thinking Skills	4
Academic Achievement	3
Critical Thinking	3
Item Analysis	3
Language Proficiency	3
Language Tests	3
Scientific Concepts	3
College Students	2
Competence	2
Computer Assisted Testing	2
Computer Science Education	2
Elementary Secondary Education	2
More ▼

Aditya Shah	1
Ajay Devmane	1
Ali Zahabi	1
Amelia Pearson	1
Andrea Vásquez	1
Anggun Resdasari Prasetyo	1
Anna V. Fisher	1
Apichat Khamboonruang	1
Aria Tsegai-Moore	1
Bambang Sumintono	1
Budi Waluyo	1
Busra Ozmen Yagiz	1
Büsra Kilinç	1
Camilo Vieira	1
Cassondra M. Eng	1
Chandralekha Singh	1
Changkyung Song	1
Charlotte Broadhurst	1
Daniel M. Bolt	1
Davis Velarde-Camaqui	1
Dina Kamber Hamzic	1
Ecenaz Alemdag	1
Elena Lieven	1
Federico Meza	1
Fridah Rotich	1
More ▼