ERIC - Search Results

Publication Date

In 2026	0
Since 2025	3
Since 2022 (last 5 years)	13

Descriptor

Computer Assisted Testing	13
Reliability	13
Validity	7
Scoring	5
Evaluation Methods	4
Foreign Countries	4
Artificial Intelligence	3
Correlation	3
Ethics	3
Grading	3
Integrity	3
Test Construction	3
Accuracy	2
Barriers	2
COVID-19	2
Cheating	2
College Faculty	2
College Students	2
Comparative Analysis	2
Computational Linguistics	2
Evaluators	2
Feedback (Response)	2
Graduate Students	2
Higher Education	2
Interrater Reliability	2
More ▼

Source

Journal of Speech, Language,…	2
Advances in Physiology…	1
American Educational Research…	1
British Educational Research…	1
Creativity Research Journal	1
International Journal for…	1
International Journal of…	1
Journal of Educational…	1
Journal of University…	1
Modern Language Journal	1
ProQuest LLC	1
Teaching in Higher Education	1
More ▼

Publication Type

Journal Articles	12
Reports - Research	11
Dissertations/Theses -…	1
Information Analyses	1
Reports - Evaluative	1

Education Level

Higher Education	8
Postsecondary Education	8
Early Childhood Education	1
Elementary Education	1
Grade 3	1
Grade 4	1
Grade 5	1
Intermediate Grades	1
Middle Schools	1
Primary Education	1

Audience

Location

India	1
North Carolina (Greensboro)	1
Oman	1
Singapore	1
South Africa	1

Laws, Policies, & Programs

Assessments and Surveys

What Works Clearinghouse Rating

Showing all 13 results Save | Export

Grading Exams Using Large Language Models: A Comparison between Human and AI Grading of Exams in Higher Education Using ChatGPT

Peer reviewed

Direct link

Jonas Flodén – British Educational Research Journal, 2025

This study compares how the generative AI (GenAI) large language model (LLM) ChatGPT performs in grading university exams compared to human teachers. Aspects investigated include consistency, large discrepancies and length of answer. Implications for higher education, including the role of teachers and ethics, are also discussed. Three…

Descriptors: College Faculty, Artificial Intelligence, Comparative Testing, Scoring

Investigating Students' Perception about LMS-Based Online Examination Practices

Peer reviewed

Direct link

Shard; Devesh Kumar; Sapna Koul – International Journal of Information and Learning Technology, 2024

Purpose: This study aims to gain insights into how students perceive online examination practices and evaluation, as well as identify the key factors that impact their intentions toward online exams. Design/methodology/approach: This empirical study conducted in India utilized an online survey method between May 24 and June 14, 2022. The data were…

Descriptors: Foreign Countries, Undergraduate Students, Graduate Students, Student Attitudes

Triangulating Learner Corpus and Online Experimental Data: Evidence from Gender Agreement and Relative Clauses in L2 Greek

Peer reviewed

Direct link

Despina Papadopoulou; Nikolaos Amvrazis; Gerakini Douka; Alexandros Tantos – Modern Language Journal, 2024

The article introduces triangulation to converge evidence from corpus and experimental data, by means of two case studies in second language (L2) learners of Greek. The first case study investigates the acquisition of gender agreement, while the second probes the development of relative clauses. In both studies, findings from the corpus are tested…

Descriptors: Greek, Phrase Structure, Second Language Learning, Second Language Instruction

Practical Randomly Selected Question Exam Design to Address Replicated and Sequential Questions in Online Examinations

Peer reviewed

Direct link

Elkhatat, Ahmed M. – International Journal for Educational Integrity, 2022

Examinations form part of the assessment processes that constitute the basis for benchmarking individual educational progress, and must consequently fulfill credibility, reliability, and transparency standards in order to promote learning outcomes and ensure academic integrity. A randomly selected question examination (RSQE) is considered to be an…

Descriptors: Integrity, Monte Carlo Methods, Credibility, Reliability

Development, Reliability, and Concurrent Validity of the American Sign Language Version of the Computerized Revised Token Test

Peer reviewed

Direct link

Emily B. Goldberg; Sheila R. Pratt; Malcolm R. McNeil; Neil Szuminsky; Kenneth DeHaan; Leslie Q. Zhen – Journal of Speech, Language, and Hearing Research, 2025

Purpose: The present study assessed the test-retest reliability of the American Sign Language (ASL) version of the Computerized Revised Token Test (CRTT-ASL) and compared the differences and similarities between ASL and English reading by Deaf and hearing users of ASL. Method: Creation of the CRTT-ASL involved filming, editing, and validating CRTT…

Descriptors: American Sign Language, Reliability, Validity, Test Construction

Classification Consistency and Results Reporting of a Digital-First Computer-Adaptive Language Proficiency Test

Direct link

Ramsey Lee Cardwell – ProQuest LLC, 2022

The emergence of digital-first assessments is prompting reconsideration of, and innovation in, aspects of psychometrics, test validation, and test use. Using the Duolingo English Test (DET) as an example, this three-paper series seeks to address issues concerning the estimation of classification consistency and the reporting of results for such…

Descriptors: Classification, Reliability, Language Proficiency, Computer Assisted Testing

Examining Human and Automated Ratings of Elementary Students' Writing Quality: A Multivariate Generalizability Theory Application

Peer reviewed

Direct link

Chen, Dandan; Hebert, Michael; Wilson, Joshua – American Educational Research Journal, 2022

We used multivariate generalizability theory to examine the reliability of hand-scoring and automated essay scoring (AES) and to identify how these scoring methods could be used in conjunction to optimize writing assessment. Students (n = 113) included subsamples of struggling writers and non-struggling writers in Grades 3-5 drawn from a larger…

Descriptors: Reliability, Scoring, Essays, Automation

The Impact of Artificial Intelligence on Online Assessment: A Preliminary Review

Peer reviewed
PDF on ERIC

Download full text

Nejdet Karadag – Journal of Educational Technology and Online Learning, 2023

The purpose of this study is to examine the impact of artificial intelligence (AI) on online assessment in the context of opportunities and threats based on the literature. To this end, 19 articles related to the AI tool ChatGPT and online assessment were analysed through rapid literature review. In the content analysis, the themes of "AI's…

Descriptors: Artificial Intelligence, Computer Assisted Testing, Natural Language Processing, Grading

Exploring the Nexus between Assessment, Quality and Social Justice: Reflections on Remote Assessment Practices

Peer reviewed

Direct link

Kershree Padayachee; M. Matimolane – Teaching in Higher Education, 2025

In the shift to Emergency Remote Teaching and Learning (ERT&L) during the COVID-19 pandemic, remote assessment and feedback became a major source of discontent and challenge for students and staff. This paper is a reflection and analysis of assessment practices during ERT&L, and our theorisation of the possibilities for shifts towards…

Descriptors: Educational Quality, Social Justice, Distance Education, Feedback (Response)

Accuracy and Reliability of Large Language Models in Assessing Learning Outcomes Achievement across Cognitive Domains

Peer reviewed

Direct link

Swapna Haresh Teckwani; Amanda Huee-Ping Wong; Nathasha Vihangi Luke; Ivan Cherh Chiet Low – Advances in Physiology Education, 2024

The advent of artificial intelligence (AI), particularly large language models (LLMs) like ChatGPT and Gemini, has significantly impacted the educational landscape, offering unique opportunities for learning and assessment. In the realm of written assessment grading, traditionally viewed as a laborious and subjective process, this study sought to…

Descriptors: Accuracy, Reliability, Computational Linguistics, Standards

Semantic Distance and the Alternate Uses Task: Recommendations for Reliable Automated Assessment of Originality

Peer reviewed

Direct link

Beaty, Roger E.; Johnson, Dan R.; Zeitlen, Daniel C.; Forthmann, Boris – Creativity Research Journal, 2022

Semantic distance is increasingly used for automated scoring of originality on divergent thinking tasks, such as the Alternate Uses Task (AUT). Despite some psychometric support for semantic distance -- including positive correlations with human creativity ratings -- additional work is needed to optimize its reliability and validity, including…

Descriptors: Semantics, Scoring, Creative Thinking, Creativity

Validation of an Automated Procedure for Calculating Core Lexicon from Transcripts

Peer reviewed

Direct link

Dalton, Sarah Grace; Stark, Brielle C.; Fromm, Davida; Apple, Kristen; MacWhinney, Brian; Rensch, Amanda; Rowedder, Madyson – Journal of Speech, Language, and Hearing Research, 2022

Purpose: The aim of this study was to advance the use of structured, monologic discourse analysis by validating an automated scoring procedure for core lexicon (CoreLex) using transcripts. Method: Forty-nine transcripts from persons with aphasia and 48 transcripts from persons with no brain injury were retrieved from the AphasiaBank database. Five…

Descriptors: Validity, Discourse Analysis, Databases, Scoring

The Impact of Online Assessment Challenges on Assessment Principles during COVID-19 in Oman

Peer reviewed
PDF on ERIC

Download full text

Al-Maqbali, Asma Hilal; Raja Hussain, Raja Maznah – Journal of University Teaching and Learning Practice, 2022

With the emergence of COVID-19, many educational pillars have been altered from conventional ways to online solutions. The educational assessment has been administered in online environments despite all encountered challenges. This descriptive study aimed to uncover the online assessment challenges that were confronted. Furthermore, it intended to…

Descriptors: Barriers, Computer Assisted Testing, COVID-19, Pandemics

Al-Maqbali, Asma Hilal	1
Alexandros Tantos	1
Amanda Huee-Ping Wong	1
Apple, Kristen	1
Beaty, Roger E.	1
Chen, Dandan	1
Dalton, Sarah Grace	1
Despina Papadopoulou	1
Devesh Kumar	1
Elkhatat, Ahmed M.	1
Emily B. Goldberg	1
Forthmann, Boris	1
Fromm, Davida	1
Gerakini Douka	1
Hebert, Michael	1
Ivan Cherh Chiet Low	1
Johnson, Dan R.	1
Jonas Flodén	1
Kenneth DeHaan	1
Kershree Padayachee	1
Leslie Q. Zhen	1
M. Matimolane	1
MacWhinney, Brian	1
Malcolm R. McNeil	1
Nathasha Vihangi Luke	1
More ▼