用因果抽象模拟语言模型如何扮演公平硬币抛掷
Heads or Tails: A Simple Example of Causal Abstractive Simulation
- 将语言模型视为角色扮演的因果系统,构建抽象模拟框架
- 展示语言模型在模拟硬币抛掷时的失败与成功案例
- 为语言模型仿真提供可验证的因果理论基础
本文通过一个简单例子说明因果抽象模拟(causal abstractive simulation)如何形式化语言模型对系统行为的模拟。研究以公平硬币抛掷为例,展示语言模型在模拟过程中可能出现的失败,并给出一个成功案例,证明在给定系统因果描述的前提下,语言模型可被形式化地证明具备模拟能力。该方法为语言模型仿真提供了连接统计基准测试与因果理论基础的桥梁,适用于语言模型模拟实践者、人工智能哲学家及因果抽象领域的数学家。该框架也为‘语言模型是角色扮演’这一观点提供了精确的操作化定义。
原文摘要 · Abstract (English)
This note illustrates how a variety of causal abstraction arXiv:1707.00819 arXiv:1812.03789, defined here as causal abstractive simulation, can be used to formalize a simple example of language model simulation. This note considers the case of simulating a fair coin toss with a language model. Examples are presented illustrating the ways language models can fail to simulate, and a success case is presented, illustrating how this formalism may be used to prove that a language model simulates some other system, given a causal description of the system. This note may be of interest to three groups. For practitioners in the growing field of language model simulation, causal abstractive simulation is a means to connect ad-hoc statistical benchmarking practices to the solid formal foundation of causality. Philosophers of AI and philosophers of mind may be interested as causal abstractive simulation gives a precise operationalization to the idea that language models are role-playing arXiv:2402.12422. Mathematicians and others working on causal abstraction may be interested to see a new application of the core ideas that yields a new variation of causal abstraction.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。