用多邻国英语测试案例,说明如何用负责任AI保障考试公平与质量。
Responsible AI for Test Equity and Quality: The Duolingo English Test as a Case Study
- 基于多邻国英语测试,建立AI伦理实践标准
- 覆盖公平性、隐私安全、可解释性等核心原则
- 适合教育评估与AI治理领域研究者参考
人工智能为考试带来效率提升,如自动化题库生成和口语写作评分,但也存在内容偏见等风险。负责任AI(RAI)旨在降低此类风险。本文以高阶英语语言测评——多邻国英语测试(Duolingo English Test, DET)为例,阐述RAI实践在保障测试质量(评分推论的合理性)与测试公平性(对所有考生的公正性)中的关键作用。文章介绍DET的RAI标准体系及其制定过程,并将其与通用型RAI原则关联。通过具体实践案例,展示如何落实有效性与可靠性、公平性、隐私与安全、透明度与问责制等伦理要求,从而确保考试的公平性与质量。
原文摘要 · Abstract (English)
Artificial intelligence (AI) creates opportunities for assessments, such as efficiencies for item generation and scoring of spoken and written responses. At the same time, it poses risks (such as bias in AI-generated item content). Responsible AI (RAI) practices aim to mitigate risks associated with AI. This chapter addresses the critical role of RAI practices in achieving test quality (appropriateness of test score inferences), and test equity (fairness to all test takers). To illustrate, the chapter presents a case study using the Duolingo English Test (DET), an AI-powered, high-stakes English language assessment. The chapter discusses the DET RAI standards, their development and their relationship to domain-agnostic RAI principles. Further, it provides examples of specific RAI practices, showing how these practices meaningfully address the ethical principles of validity and reliability, fairness, privacy and security, and transparency and accountability standards to ensure test equity and quality.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。