为大模型研究提供可落地的伦理指南,助你避开常见陷阱。
The Only Way is Ethics: A Guide to Ethical Research with Large Language Models
- 梳理百篇文献,提炼研究各阶段的伦理行动清单
- 给出具体可操作的'该做'与'不该做'建议
- 适合所有从事大模型研究的工程师和学者参考
当前已有大量关于大语言模型(LLMs)伦理问题的研究:批判评估工具与危害;提出辅助构思的工具包;讨论对工作者的影响;探讨隐私与安全相关的立法等。然而,尚无将这些资源整合成针对LLMs的实用指南的工作。本文旨在实现这一目标,推出开放且持续更新的《LLM伦理白皮书》,供自然语言处理从业者及评估他人工作伦理影响者使用。目标是将伦理文献转化为计算机科学家可执行的具体建议与思考启发,包含清晰的初步步骤。本白皮书从全面文献综述中提炼出明确的‘应做’与‘不应做’清单,并推荐有助于伦理工作的工具包。感兴趣的读者可查阅完整版《LLM伦理白皮书》,其中在项目生命周期的每个阶段都提供了简明的伦理考量说明,并附有数百篇参考文献。本文可视为开展大模型伦理研究的便携指南。
原文摘要 · Abstract (English)
There is a significant body of work looking at the ethical considerations of large language models (LLMs): critiquing tools to measure performance and harms; proposing toolkits to aid in ideation; discussing the risks to workers; considering legislation around privacy and security etc. As yet there is no work that integrates these resources into a single practical guide that focuses on LLMs; we attempt this ambitious goal. We introduce 'LLM Ethics Whitepaper', which we provide as an open and living resource for NLP practitioners, and those tasked with evaluating the ethical implications of others' work. Our goal is to translate ethics literature into concrete recommendations and provocations for thinking with clear first steps, aimed at computer scientists. 'LLM Ethics Whitepaper' distils a thorough literature review into clear Do's and Don'ts, which we present also in this paper. We likewise identify useful toolkits to support ethical work. We refer the interested reader to the full LLM Ethics Whitepaper, which provides a succinct discussion of ethical considerations at each stage in a project lifecycle, as well as citations for the hundreds of papers from which we drew our recommendations. The present paper can be thought of as a pocket guide to conducting ethical research with LLMs.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。