A framework for evaluation of large language models in essay assessment: Reliability, alignment, and causal reasoning
Publication date: June 2026 Source: Computers and Education: Artificial Intelligence, Volume 10 Author(s): Tongxi Liu, Luyao Ye, Wei Yan
Record details
Published: 14 July 2026
Source: Computers and Education: Artificial Intelligence
Category: Research
Topics: Safety & alignment · Children & education
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
LLM sentiment quantification reveals selective alignment with human course-evaluation raters
Computers and Education: Artificial Intelligence · 14 July 2026
Fine-Tuning, Retrieval-Augmented Generation, and Hybrid Large Language Models for Postoperative Decision Support: Comparative Analysis
JMIR (Journal of Medical Internet Research) · 14 July 2026
When AI Takes the Couch: Psychometric Jailbreaks Reveal Internal Conflict in Frontier Models
arXiv cs.CY · 21 July 2026
D2PO: Optimizing Diffusion Samplers via Dynamic Preference
arXiv · 7 July 2026
Covert Trait Propagation Is Representation Alignment: Mechanistic Evidence from Hidden-Channel Distillation
arXiv · 5 July 2026
X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment
arXiv cs.LG · 23 July 2026
How to cite this record
ethics.ai (14 July 2026), “A framework for evaluation of large language models in essay assessment: Reliability, alignment, and causal reasoning,” evidence record 2384, https://ethics.ai/record/2384 (originally published by Computers and Education: Artificial Intelligence).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.