ERIC - Search Results

Publication Date

In 2026	0
Since 2025	0
Since 2022 (last 5 years)	0
Since 2017 (last 10 years)	9
Since 2007 (last 20 years)	37

Descriptor

Language Tests	44
Statistical Analysis	44
Second Language Learning	32
English (Second Language)	28
Foreign Countries	16
Comparative Analysis	15
Language Proficiency	14
Scores	13
Testing	12
Oral Language	10
Correlation	9
Test Items	9
College Students	7
Evaluators	7
Scoring	7
Second Language Instruction	7
Difficulty Level	6
Item Response Theory	6
Native Speakers	6
Reading Comprehension	6
Test Validity	6
Computer Assisted Testing	5
Elementary School Students	5
Item Analysis	5
Models	5
More ▼

Source

Language Testing

Publication Type

Journal Articles	44
Reports - Research	29
Reports - Evaluative	11
Tests/Questionnaires	4
Information Analyses	3
Reports - Descriptive	2
Guides - Non-Classroom	1
Opinion Papers	1
Speeches/Meeting Papers	1

Education Level

Higher Education	12
Postsecondary Education	8
Elementary Education	4
Secondary Education	4
Elementary Secondary Education	1
Grade 5	1
High Schools	1
Intermediate Grades	1

Audience

Location

Japan	5
Australia	3
United Kingdom	2
Canada	1
Chile	1
China	1
Colombia	1
Ecuador	1
Germany	1
Israel	1
Malaysia	1
Poland	1
Russia	1
Sweden	1
More ▼

Laws, Policies, & Programs

Assessments and Surveys

Test of English as a Foreign…	6
Early Childhood Longitudinal…	1
International English…	1
Michigan Test of English…	1

What Works Clearinghouse Rating

Showing 1 to 15 of 44 results Save | Export

A Systematic Review of Methods for Evaluating Rating Quality in Language Assessment

Peer reviewed

Direct link

Wind, Stefanie A.; Peterson, Meghan E. – Language Testing, 2018

The use of assessments that require rater judgment (i.e., rater-mediated assessments) has become increasingly popular in high-stakes language assessments worldwide. Using a systematic literature review, the purpose of this study is to identify and explore the dominant methods for evaluating rating quality within the context of research on…

Descriptors: Language Tests, Evaluators, Evaluation Methods, Interrater Reliability

Setting Cut Scores on an EFL Placement Test Using the Prototype Group Method: A Receiver Operating Characteristic (ROC) Analysis

Peer reviewed

Direct link

Eckes, Thomas – Language Testing, 2017

This paper presents an approach to standard setting that combines the prototype group method (PGM; Eckes, 2012) with a receiver operating characteristic (ROC) analysis. The combined PGM-ROC approach is applied to setting cut scores on a placement test of English as a foreign language (EFL). To implement the PGM, experts first named learners whom…

Descriptors: English (Second Language), Language Tests, Cutting Scores, Standard Setting (Scoring)

The Selection of Cognitive Diagnostic Models for a Reading Comprehension Test

Peer reviewed

Direct link

Li, Hongli; Hunter, C. Vincent; Lei, Pui-Wa – Language Testing, 2016

Cognitive diagnostic models (CDMs) have great promise for providing diagnostic information to aid learning and instruction, and a large number of CDMs have been proposed. However, the assumptions and performances of different CDMs and their applications in regard to reading comprehension tests are not fully understood. In the present study, we…

Descriptors: Reading Comprehension, Reading Tests, Models, Comparative Analysis

Investigating the Construct Measured by Banked Gap-Fill Items: Evidence from Eye-Tracking

Peer reviewed

Direct link

McCray, Gareth; Brunfaut, Tineke – Language Testing, 2018

This study investigates test-takers' processing while completing banked gap-fill tasks, designed to test reading proficiency, in order to test theoretically based expectations about the variation in cognitive processes of test-takers across levels of performance. Twenty-eight test-takers' eye traces on 24 banked gap-fill items (on six tasks) were…

Descriptors: Language Tests, Test Items, Item Analysis, Eye Movements

The Influence of Training and Experience on Rater Performance in Scoring Spoken Language

Peer reviewed

Direct link

Davis, Larry – Language Testing, 2016

Two factors were investigated that are thought to contribute to consistency in rater scoring judgments: rater training and experience in scoring. Also considered were the relative effects of scoring rubrics and exemplars on rater performance. Experienced teachers of English (N = 20) scored recorded responses from the TOEFL iBT speaking test prior…

Descriptors: Evaluators, Oral Language, Scores, Language Tests

Young Learners' Response Processes When Taking Computerized Tasks for Speaking Assessment

Peer reviewed

Direct link

Lee, Shinhye; Winke, Paula – Language Testing, 2018

We investigated how young language learners process their responses on and perceive a computer-mediated, timed speaking test. Twenty 8-, 9-, and 10-year-old non-native English-speaking children (NNSs) and eight same-aged, native English-speaking children (NSs) completed seven computerized sample TOEFL® Primary™ speaking test tasks. We investigated…

Descriptors: Elementary School Students, Second Language Learning, Responses, Computer Assisted Testing

Task and Rater Effects in L2 Speaking and Writing: A Synthesis of Generalizability Studies

Peer reviewed

Direct link

In'nami, Yo; Koizumi, Rie – Language Testing, 2016

We addressed Deville and Chalhoub-Deville's (2006), Schoonen's (2012), and Xi and Mollaun's (2006) call for research into the contextual features that are considered related to person-by-task interactions in the framework of generalizability theory in two ways. First, we quantitatively synthesized the generalizability studies to determine the…

Descriptors: Evaluators, Second Language Learning, Writing Skills, Oral Language

Evaluating Different Standard-Setting Methods in an ESL Placement Testing Context

Peer reviewed

Direct link

Shin, Sun-Young; Lidster, Ryan – Language Testing, 2017

In language programs, it is crucial to place incoming students into appropriate levels to ensure that course curriculum and materials are well targeted to their learning needs. Deciding how and where to set cutscores on placement tests is thus of central importance to programs, but previous studies in educational measurement disagree as to which…

Descriptors: Language Tests, English (Second Language), Standard Setting (Scoring), Student Placement

Elicited Imitation as a Measure of Second Language Proficiency: A Narrative Review and Meta-Analysis

Peer reviewed

Direct link

Yan, Xun; Maeda, Yukiko; Lv, Jing; Ginther, April – Language Testing, 2016

Elicited imitation (EI) has been widely used to examine second language (L2) proficiency and development and was an especially popular method in the 1970s and early 1980s. However, as the field embraced more communicative approaches to both instruction and assessment, the use of EI diminished, and the construct-related validity of EI scores as a…

Descriptors: Second Language Learning, Language Proficiency, Meta Analysis, Effect Size

Assessing Syntactic Sophistication in L2 Writing: A Usage-Based Approach

Peer reviewed

Direct link

Kyle, Kristopher; Crossley, Scott – Language Testing, 2017

Over the past 45 years, the construct of syntactic sophistication has been assessed in L2 writing using what Bulté and Housen (2012) refer to as absolute complexity (Lu, 2011; Ortega, 2003; Wolfe-Quintero, Inagaki, & Kim, 1998). However, it has been argued that making inferences about learners based on absolute complexity indices (e.g., mean…

Descriptors: Syntax, Verbs, Second Language Learning, Word Frequency

Using Corpus Linguistics to Examine the Extrapolation Inference in the Validity Argument for a High-Stakes Speaking Assessment

Peer reviewed

Direct link

LaFlair, Geoffrey T.; Staples, Shelley – Language Testing, 2017

Investigations of the validity of a number of high-stakes language assessments are conducted using an argument-based approach, which requires evidence for inferences that are critical to score interpretation (Chapelle, Enright, & Jamieson, 2008b; Kane, 2013). The current study investigates the extrapolation inference for a high-stakes test of…

Descriptors: Computational Linguistics, Language Tests, Test Validity, Inferences

Determining Cloze Item Difficulty from Item and Passage Characteristics across Different Learner Backgrounds

Peer reviewed

Direct link

Trace, Jonathan; Brown, James Dean; Janssen, Gerriet; Kozhevnikova, Liudmila – Language Testing, 2017

Cloze tests have been the subject of numerous studies regarding their function and use in both first language and second language contexts (e.g., Jonz & Oller, 1994; Watanabe & Koyama, 2008). From a validity standpoint, one area of investigation has been the extent to which cloze tests measure reading ability beyond the sentence level.…

Descriptors: Cloze Procedure, Language Tests, Test Items, Item Analysis

Construct Validity in TOEFL iBT Speaking Tasks: Insights from Natural Language Processing

Peer reviewed

Direct link

Kyle, Kristopher; Crossley, Scott A.; McNamara, Danielle S. – Language Testing, 2016

This study explores the construct validity of speaking tasks included in the TOEFL iBT (e.g., integrated and independent speaking tasks). Specifically, advanced natural language processing (NLP) tools, MANOVA difference statistics, and discriminant function analyses (DFA) are used to assess the degree to which and in what ways responses to these…

Descriptors: Construct Validity, Natural Language Processing, Speech Skills, Speech Acts

Lexical Difficulty--Using Elicited Imitation to Study Child L2

Peer reviewed

Direct link

Campfield, Dorota E. – Language Testing, 2017

This paper reports a post-hoc analysis of the influence of lexical difficulty of cue sentences on performance in an elicited imitation (EI) task to assess oral production skills for 645 child L2 English learners in instructional settings. This formed part of a large-scale investigation into effectiveness of foreign language teaching in Polish…

Descriptors: Difficulty Level, Second Language Learning, Second Language Instruction, Elementary School Students

How Do Utterance Measures Predict Raters' Perceptions of Fluency in French as a Second Language?

Peer reviewed

Direct link

Préfontaine, Yvonne; Kormos, Judit; Johnson, Daniel Ezra – Language Testing, 2016

While the research literature on second language (L2) fluency is replete with descriptions of fluency and its influence with regard to English as an additional language, little is known about what fluency features influence judgments of fluency in L2 French. This study reports the results of an investigation that analyzed the relationship between…

Descriptors: Prediction, French, Second Language Learning, Evaluators

Previous Page | Next Page »

Pages: 1 | 2 | 3

Bachman, Lyle F.	2
Bae, Jungok	2
Crossley, Scott A.	2
Kyle, Kristopher	2
McNamara, Danielle S.	2
Allalouf, Avi	1
Alvarez, Marta E.	1
Batty, Aaron Olaf	1
Bax, Stephen	1
Boo, Jaeyool	1
Brown, James Dean	1
Brunfaut, Tineke	1
Campfield, Dorota E.	1
Chalhoub-Deville, Micheline	1
Chen, Fang	1
Choi, Inn-Chull	1
Crossley, Scott	1
Davies, Alan	1
Davis, Larry	1
Eckes, Thomas	1
Feng, Ying	1
Filipi, Anna	1
Garras, John	1
Ginther, April	1
More ▼