LLMs cannot be relied on to give accurate indication of student performance, researchers find, as universities explore ways to relieve pressure on graders. Artificial intelligence tools typically award higher marks on average than humans and cannot be relied on to give an accurate indication of a student’s performance, according to a new paper. With AI increasingly being explored by universities as part of the marking process to relieve pressure on time-strapped academics, the research, published in the journal Assessment & Evaluation in Higher Education, found that generative AI tools such as ChatGPT “do not reliably reproduce human judgement in the marking of extended written work.”
No comments:
Post a Comment