From Inside Higher Education: “[R]esearch, published in the journal Assessment & Evaluation in Higher Education, found that generative AI tools such as ChatGPT `do not reliably reproduce human judgement in the marking of extended written work.’ . . . In all but one case [out of 50], the AI models typically returned higher average marks than humans. In one case the difference between an AI grade and a human one was 40 points for an essay where the top score was 100.”
Might professors be willing defer to AI in the grading of exams? An interesting mix of responses could be expected to such a proposal.



