ERIC - Search Results

Publication Date

In 2025	2
Since 2024	3
Since 2021 (last 5 years)	7
Since 2016 (last 10 years)	34
Since 2006 (last 20 years)	77

Descriptor

Standard Setting (Scoring)	211
Cutting Scores	121
Standards	55
Test Items	55
Higher Education	41
Licensing Examinations…	40
Minimum Competency Testing	37
Interrater Reliability	35
Comparative Analysis	33
Difficulty Level	32
Elementary Secondary Education	31
Evaluation Methods	29
Foreign Countries	29
Scoring	28
Judges	27
Scores	26
Evaluators	25
Academic Standards	24
Test Validity	23
Criterion Referenced Tests	21
Item Response Theory	21
Mathematics Tests	20
Testing Programs	18
Certification	17
Error of Measurement	17
More ▼

Publication Type

Reports - Research	211
Journal Articles	117
Speeches/Meeting Papers	71
Tests/Questionnaires	8
Reports - Evaluative	4
Numerical/Quantitative Data	2
Collected Works - General	1
Collected Works - Proceedings	1
Information Analyses	1
Legal/Legislative/Regulatory…	1
Opinion Papers	1
More ▼

Education Level

Higher Education	18
Postsecondary Education	14
Secondary Education	12
Elementary Education	10
Middle Schools	8
Elementary Secondary Education	6
Intermediate Grades	6
Junior High Schools	6
Grade 5	4
Early Childhood Education	2
Grade 3	2
Grade 4	2
Grade 6	2
Grade 7	2
Grade 8	2
High Schools	2
Primary Education	2
Grade 11	1
Grade 2	1
More ▼

Audience

Researchers	23
Policymakers	1
Teachers	1

Location

Tennessee	6
Canada	5
Australia	4
California	4
United Kingdom	4
Kansas	3
New Jersey	3
North Carolina	3
Germany	2
Illinois	2
Indiana	2
Massachusetts	2
Netherlands	2
Nevada	2
Taiwan	2
Turkey	2
Washington	2
Africa	1
Arizona	1
Arkansas	1
Colorado	1
Delaware	1
France	1
Georgia	1
Idaho	1
More ▼

Laws, Policies, & Programs

No Child Left Behind Act 2001	2
Comprehensive Education…	1
Education Consolidation…	1

Assessments and Surveys

National Teacher Examinations	14
Alabama High School…	4
National Assessment of…	3
Praxis Series	2
Test of English as a Foreign…	2
United States Medical…	2
edTPA (Teacher Performance…	2
Advanced Placement…	1
College Board Achievement…	1
Massachusetts Comprehensive…	1
Pre Professional Skills Tests	1
Wechsler Adult Intelligence…	1
More ▼

What Works Clearinghouse Rating

Showing 1 to 15 of 211 results Save | Export

Embedding Embedded Standard Setting: An Application of Cross-Classified Item Response Theory. CRESST Report 876

Download full text

Yun-Kyung Kim; Li Cai – National Center for Research on Evaluation, Standards, and Student Testing (CRESST), 2025

This paper introduces an application of cross-classified item response theory (IRT) modeling to an assessment utilizing the embedded standard setting (ESS) method (Lewis & Cook). The cross-classified IRT model is used to treat both item and person effects as random, where the item effects are regressed on the target performance levels (target…

Descriptors: Standard Setting (Scoring), Item Response Theory, Test Items, Difficulty Level

Setting and Validating Multiple Standards on a Multistage-Adaptive Test

Peer reviewed

Direct link

Lewis, Jennifer; Lim, Hwanggyu; Padellaro, Frank; Sireci, Stephen G.; Zenisky, April L. – Educational Measurement: Issues and Practice, 2022

Setting cut scores on (MSTs) is difficult, particularly when the test spans several grade levels, and the selection of items from MST panels must reflect the operational test specifications. In this study, we describe, illustrate, and evaluate three methods for mapping panelists' Angoff ratings into cut scores on the scale underlying an MST. The…

Descriptors: Cutting Scores, Adaptive Testing, Test Items, Item Analysis

The Choice of Response Probability in Bookmark Standard Setting: An Experimental Study

Peer reviewed

Direct link

Baldwin, Peter; Margolis, Melissa J.; Clauser, Brian E.; Mee, Janet; Winward, Marcia – Educational Measurement: Issues and Practice, 2020

Evidence of the internal consistency of standard-setting judgments is a critical part of the validity argument for tests used to make classification decisions. The bookmark standard-setting procedure is a popular approach to establishing performance standards, but there is relatively little research that reflects on the internal consistency of the…

Descriptors: Standard Setting (Scoring), Probability, Cutting Scores, Evaluation Methods

A Critical Look into the Beuk Standard-Setting Method

Peer reviewed

Direct link

Wyse, Adam E. – Educational Measurement: Issues and Practice, 2020

One commonly used compromise standard-setting method is the Beuk (1984) method. A key assumption of the Beuk method is that the emphasis given to the pass rate and the percent correct ratings should be proportional to the extent that the panelists agree on their ratings. However, whether the slope of Beuk line reflects the emphasis that panelists…

Descriptors: Standard Setting (Scoring), Cutting Scores, Weighted Scores, Evaluation Methods

Comparing Cut Scores from the Angoff Method and Two Variations of the Hofstee and Beuk Methods

Peer reviewed

Direct link

Wyse, Adam E. – Applied Measurement in Education, 2020

This article compares cut scores from two variations of the Hofstee and Beuk methods, which determine cut scores by resolving inconsistencies in panelists' judgments about cut scores and pass rates, with the Angoff method. The first variation uses responses to the Hofstee and Beuk percentage correct and pass rate questions to calculate cut scores.…

Descriptors: Cutting Scores, Evaluation Methods, Standard Setting (Scoring), Equations (Mathematics)

Using Diagnostic Profiles to Describe Borderline Performance in Standard Setting

Peer reviewed

Direct link

Skaggs, Gary; Hein, Serge F.; Wilkins, Jesse L. M. – Educational Measurement: Issues and Practice, 2020

In test-centered standard-setting methods, borderline performance can be represented by many different profiles of strengths and weaknesses. As a result, asking panelists to estimate item or test performance for a hypothetical group study of borderline examinees, or a typical borderline examinee, may be an extremely difficult task and one that can…

Descriptors: Standard Setting (Scoring), Cutting Scores, Testing Problems, Profiles

The Riddle Knowledge Inference Test (R-Kit)

Peer reviewed

Direct link

Nicolas Rochat; Laurent Lima; Pascal Bressoux – Journal of Psychoeducational Assessment, 2025

Inference is considered an important factor in comprehension models and has been described as a causal factor in predicting comprehension. To date, specific tests for inference are rare and often rely on specific thematic texts. This reliance on thematic inference may raise some concerns as inference is related to prior text-specific knowledge.…

Descriptors: Inferences, Reading Comprehension, Reading Tests, Test Reliability

Making the Grade with Recreational Therapy Accreditation: Comparing the NCTRC Pass Rates of CAAHEP/CARTE Accredited Programs to National Averages

Peer reviewed

Direct link

David Loy; Rhonda Nelson; Jared Allsop; Carol Johnston – Schole: A Journal of Leisure Studies and Recreation Education, 2024

Accreditation is a critical process in maintaining standards of consistency and excellence in the academic preparation of students for their chosen profession. While academic programs, professional associations, and credentialing organizations all recognize the importance of programmatic accreditation in recreational therapy professional…

Descriptors: Therapeutic Recreation, Accreditation (Institutions), Scores, Tests

Examining the Impact of a Consensus Approach to Content Alignment Studies

Peer reviewed
PDF on ERIC

Download full text

Russell, Michael; Moncaleano, Sebastian – Practical Assessment, Research & Evaluation, 2020

Although both content alignment and standard-setting procedures rely on content-expert panel judgements, only the latter employs discussion among panel members. This study employed a modified form of the Webb methodology to examine content alignment for twelve tests administered as part of the Massachusetts Comprehensive Assessment System (MCAS).…

Descriptors: Test Content, Test Items, Discussion, Test Validity

Mapping "TOEFL® Essentials"™ Test Scores to the Canadian Language Benchmarks. "TOEFL"® Research Report. TOEFL-RR-100. ETS Research Report No. RR-22-16

Peer reviewed
PDF on ERIC

Download full text

Papageorgiou, Spiros; Davis, Larry; Ohta, Renka; Gomez, Pablo Garcia – ETS Research Report Series, 2022

In this research report, we describe a study to map the scores of the "TOEFL® Essentials"™ test to the Canadian Language Benchmarks (CLB). The TOEFL Essentials test is a four-skills assessment of foundational English language skills and communication abilities in academic and general (daily life) contexts. At the time of writing this…

Descriptors: Foreign Countries, Language Tests, English (Second Language), Second Language Learning

It's Not Just Angoff: Misperceptions of Hard and Easy Items in Bookmark-Type Ratings

Peer reviewed

Direct link

Wyse, Adam E.; Babcock, Ben – Educational Measurement: Issues and Practice, 2020

A common belief is that the Bookmark method is a cognitively simpler standard-setting method than the modified Angoff method. However, a limited amount of research has investigated panelist's ability to perform well the Bookmark method, and whether some of the challenges panelists face with the Angoff method may also be present in the Bookmark…

Descriptors: Standard Setting (Scoring), Evaluation Methods, Testing Problems, Test Items

Applicability of Two Standard Setting Methods for Enhancing the Reporting of Assessment Results within the South African Education Context

Peer reviewed
PDF on ERIC

Download full text

Moloi, Qetelo; Kanjee, Anil – South African Journal of Education, 2021

The study reported on here contributes to the growing body of knowledge on the use of standard setting methods for improving the reporting and utility value of assessment results in South Africa as well as for addressing the conceptual shortcomings of the Curriculum and Assessment Policy Statement (CAPS) reporting framework. Using data from the…

Descriptors: Foreign Countries, Standard Setting (Scoring), Student Evaluation, Elementary School Students

An Evaluation of Pass/Fail Decisions through Norm- and Criterion-Referenced Assessments

Peer reviewed
PDF on ERIC

Download full text

Cuhadar, Ismail; Gelbal, Selahattin – International Journal of Assessment Tools in Education, 2021

The institutions in education use various assessment methods to decide on the proficiency levels of students in a particular construct. This study investigated whether the decisions differed based on the type of assessment: norm-and criterion-referenced assessment. An achievement test with 20 multiple-choice items was administered to 107 students…

Descriptors: Norm Referenced Tests, Criterion Referenced Tests, Decision Making, Achievement Tests

Equating Angoff Standard-Setting Ratings with the Rasch Model

Peer reviewed

Direct link

Wyse, Adam E. – Measurement: Interdisciplinary Research and Perspectives, 2018

A key part of determining cut-scores when performing Angoff standard setting is utilizing equating methods to place standard-setting ratings onto the scale used to report scores to examinees. This article describes three equating methods that can be employed to place Angoff ratings onto the scale used to report scores to examinees when applying…

Descriptors: Standard Setting (Scoring), Equated Scores, Probability, Regression (Statistics)

Comparison of Passing Scores Determined by the Angoff Method in Different Item Samples

Peer reviewed
PDF on ERIC

Download full text

Kara, Hakan; Cetin, Sevda – International Journal of Assessment Tools in Education, 2020

In this study, the efficiency of various random sampling methods to reduce the number of items rated by judges in an Angoff standard-setting study was examined and the methods were compared with each other. Firstly, the full-length test was formed by combining Placement Test 2012 and 2013 mathematics subsets. After then, simple random sampling…

Descriptors: Cutting Scores, Standard Setting (Scoring), Sampling, Error of Measurement

Previous Page | Next Page »

Pages: 1 | 2 | 3 | 4 | 5 | 6 | 7 | 8 | 9 | 10 | 11 | ... | 15

Applied Measurement in…	19
Journal of Educational…	19
Educational Measurement:…	12
Educational and Psychological…	11
Evaluation and the Health…	4
Practical Assessment,…	4
Assessment & Evaluation in…	3
Educational Assessment	2
Human Resources Research…	2
International Journal of…	2
International Journal of…	2
Language Assessment Quarterly	2
Language Testing	2
Measurement:…	2
National Center for Research…	2
Online Submission	2
Studies in Educational…	2
Academic Medicine	1
Advances in Health Sciences…	1
Assessment	1
Assessment in Education:…	1
Canadian Modern Language…	1
Contemporary Issues in Reading	1
ETS Research Report Series	1
Educational Administration…	1
More ▼

Plake, Barbara S.	15
Wyse, Adam E.	12
Impara, James C.	9
Clauser, Brian E.	8
Jaeger, Richard M.	8
Margolis, Melissa J.	8
Busch, John Christian	6
Norcini, John J.	6
Bowman, Harry L.	5
Giraud, Gerald	5
Mee, Janet	5
Livingston, Samuel A.	4
Shulruf, Boaz	4
Chang, Lei	3
Clauser, Jerome C.	3
Ferdous, Abdullah A.	3
Halpin, Gerald	3
Hambleton, Ronald K.	3
Kannan, Priya	3
Petry, John R.	3
Sireci, Stephen G.	3
Tannenbaum, Richard J.	3
Babcock, Ben	2
Buckendahl, Chad W.	2
More ▼