arXiv:2409.15461cs.AIcs.CL2024-09被引 4

用多专家协作生成符合人文教育规范的对话数据。

RAM2C: A Liberal Arts Educational Chatbot based on Retrieval-augmented Multi-role Multi-expert Collaboration

  • 构建三类知识库,驱动多角色LLM协作生成教学对话。
  • 在语文阅读教学中显著提升个性化与伦理安全性。
  • 适合教育AI研发者和需合规对话生成的场景。

近年来,大量研究将大语言模型(LLMs)应用于教育对话。尤其在人文学科对话中,教育者需兼顾人性化沟通、教学专长与安全伦理(HTS),而不仅仅是学科知识。然而,从真实世界收集大量符合HTS标准的教学对话作为训练语料成本高昂,现有LLMs在教学对话中的输出仍难达人类水平。为此,我们设计了检索增强的多角色多专家协作框架(RAM2C),用于自动生成此类对话数据。首先,建立涵盖教学技能、心理学与安全伦理三方面的HTS引导知识库;随后,将检索增强的LLMs组织为具有不同角色的多专家组,协同生成符合HTS标准的教育对话数据集,并在此基础上微调LLMs。实证评估表明,经RAM2C增强的LLMs在中文阅读教学中表现优异,提供更个性化且伦理安全的教学回应,验证了该框架的实用性和高质量。实验代码已公开于:https://github.com/ram2c/ram2c。

原文摘要 · Abstract (English)

Recently, many studies focus on utilizing large language models (LLMs) into educational dialogues. Especially, within liberal arts dialogues, educators must balance \textbf{H}umanized communication, \textbf{T}eaching expertise, and \textbf{S}afety-ethics (\textbf{HTS}), besides the subject knowledge itself. However, due to collecting massive amounts of HTS-compliant teaching dialogues from real world as training corpus is expensive, the outputs of existing LLMs in teaching dialogues fall short of human standards. To address this, we design a Retrieval-augmented Multi-role Multi-expert Collaboration (RAM2C) framework to automatically generate such dialogues data. Specifically, we first establish HTS-guided knowledge bases, encompassing three domain knowledge in teaching skills, psychology, and safety ethics. Then, RAM2C organizes LLMs, which are retrieval-augmented by the above different knowledge bases, into multi-experts groups with distinct roles to generate the HTS-compliant educational dialogues dataset. We then fine-tuned the LLMs using this dataset. Empirical evaluations indicate that RM2C-empowered LLMs excel in Chinese reading teaching, offering more personalized, and ethically safe teaching response, demonstrating RAM2C's practicality and high quality. We release the experiments at \hyperlink{https://github.com/ram2c/ram2c}{https://github.com/ram2c/ram2c}.

教育AI多专家协作伦理安全对话生成

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。