arXiv:2603.14443cs.CL2026-03

用计算方法分析波斯诗人百年语音风格,发现声音特征受时代与体裁影响。

Echoes Across Centuries: Phonetic Signatures of Persian Poets

  • 基于110万行诗构建语音指标,控制格律和体裁后分析诗人差异
  • 83位诗人展现稳定语音模式,如高响度抒情、硬音修辞等类型
  • 适合文学史、计算语言学及跨学科研究者参考

本研究将波斯诗歌的语音纹理视为文学历史现象,而非韵律产物或分类特征。分析基于包含31,988首诗、共1,116,306行的大型语料库,限定五种主要古典格律以实现可控比较。每行转换为音素表示,采用六项语音指标:硬度、响度、嘶音度、元音比例、音素熵与辅音簇比例。统计模型在控制格律、诗体与行长的前提下估计诗人层面差异。结果显示,尽管格律与诗体解释了大量语音变异,但诗人间系统性差异依然存在。波斯诗歌语音表现为共享韵律结构内的条件化变异,而非纯个人风格或简单韵律残留。多维风格图谱揭示若干典型语音特征,包括高响度抒情风格、以硬度驱动的修辞或史诗风格、嘶音主导的神秘主义轮廓以及高熵复杂质感。历史分析表明,语音分布随世纪演变,反映体裁兴衰、文坛制度与表演语境变迁,而非突变。本研究建立波斯诗歌语音分析的语料库级框架,证明计算语音学可助力文学历史解读,同时尊重塑造波斯诗体的形式结构。

原文摘要 · Abstract (English)

This study examines phonetic texture in Persian poetry as a literary-historical phenomenon rather than a by-product of meter or a feature used only for classification. The analysis draws on a large corpus of 1,116,306 mesras from 31,988 poems written by 83 poets, restricted to five major classical meters to enable controlled comparison. Each line is converted into a grapheme-to-phoneme representation and analyzed using six phonetic metrics: hardness, sonority, sibilance, vowel ratio, phoneme entropy, and consonant-cluster ratio. Statistical models estimate poet-level differences while controlling for meter, poetic form, and line length. The results show that although meter and form explain a substantial portion of phonetic variation, they do not eliminate systematic differences between poets. Persian poetic sound therefore appears as conditioned variation within shared prosodic structures rather than as either purely individual style or simple metrical residue. A multidimensional stylistic map reveals several recurrent phonetic profiles, including high-sonority lyric styles, hardness-driven rhetorical or epic styles, sibilant mystical contours, and high-entropy complex textures. Historical analysis indicates that phonetic distributions shift across centuries, reflecting changes in genre prominence, literary institutions, and performance contexts rather than abrupt stylistic breaks. The study establishes a corpus-scale framework for phonetic analysis in Persian poetry and demonstrates how computational phonetics can contribute to literary-historical interpretation while remaining attentive to the formal structures that shape Persian verse.

语音分析文学史计算人文波斯诗歌

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。