用数字方法生成文本谱系图,推动文献研究从结论转向探索。
A digital perspective on the role of a stemma in material-philological transmission studies
- 通过自动化工具将文本数据转为树状关系图,快速生成谱系分析起点。
- 以古诺斯语《霍隆姆达尔萨迦》为例,验证谱系图可揭示未解文本问题。
- 开源数据与脚本支持复现,适合文献学与数字人文研究者使用。
基于数字人文学科发展和学术工作流自动化趋势,本文探讨数字方法对文本传统研究的影响。认为计算机生成的谱系图(stemma codicum)应作为研究起点而非最终成果。以古诺斯语《霍隆姆达尔萨迦》为案例,证明谱系图能开启新研究路径,解决以往无法回答的问题。文章附带用于构建谱系图的数据集及两个定制Python脚本,可将依据TEI标准编码的XML文本数据转换为PHYLIP包所需的输入格式,用于生成无根文本关系树。
原文摘要 · Abstract (English)
Taking its point of departure in the recent developments in the field of digital humanities and the increasing automatisation of scholarly workflows, this study explores the implications of digital approaches to textual traditions for the broader field of textual scholarship. It argues that the relative simplicity of creating computergenerated stemmas allows us to view the stemma codicum as a research tool rather than the final product of our scholarly investigation. Using the Old Norse saga of Hrómundur as a case study, this article demonstrates that stemmas can serve as a starting point for exploring textual traditions further. In doing so, they enable us to address research questions that otherwise remain unanswered. The article is accompanied by datasets used to generate stemmas for the Hrómundar saga tradition as well as two custom Python scripts. The scripts are designed to convert XML-based textual data, encoded according to the TEI Guidelines, into the input format used for the analysis in the PHYLIP package to generate unrooted trees of relationships between texts.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。