用大模型自动评卷并解释评分理由,让评估更透明高效。
An Automated Explainable Educational Assessment System Built on LLMs
- 基于大模型自动生成评分与解释性理由
- 支持交互式可视化,便于验证评分逻辑
- 适合教育研究者与教师快速评估学生作答
本演示介绍 AERA Chat,一个用于学生作答交互式与可视化评估的自动化可解释教育评估系统。该系统利用大语言模型(LLMs)生成自动评分及评分理由,解决自动化教育评估中解释性不足和标注成本高的问题。用户可输入题目与学生答案,系统为教育工作者与研究人员提供评估准确性和大模型生成理由质量的洞察。此外,系统还提供高级可视化与稳健的评估工具,提升教育评估的可用性,并促进评分理由的有效验证。演示视频见 https://youtu.be/qUSjz-sxlBc。
原文摘要 · Abstract (English)
In this demo, we present AERA Chat, an automated and explainable educational assessment system designed for interactive and visual evaluations of student responses. This system leverages large language models (LLMs) to generate automated marking and rationale explanations, addressing the challenge of limited explainability in automated educational assessment and the high costs associated with annotation. Our system allows users to input questions and student answers, providing educators and researchers with insights into assessment accuracy and the quality of LLM-assessed rationales. Additionally, it offers advanced visualization and robust evaluation tools, enhancing the usability for educational assessment and facilitating efficient rationale verification. Our demo video can be found at https://youtu.be/qUSjz-sxlBc.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。