arXiv:2604.03265cs.GLcs.CL2026-04

首篇用印度语撰写的计算机科学论文,探索技术术语本土化路径

On the First Computer Science Research Paper in an Indian Language and the Future of Science in Indian Languages

  • 用梵语语法体系构建泰卢固语技术术语,解决科研表达难题
  • 开发泰卢固语XeLaTeX模板,突破印度语数学排版瓶颈
  • 为印地语等亿级母语者科学写作提供可复制的范式

本文描述了撰写首篇完全以印度语言——泰卢固语(约1亿使用者)表达的现代计算机科学原创研究论文的经历。该论文聚焦分布式计算领域,提出一种基于认识逻辑证明多处理器算法下界的技巧。主要挑战在于为算法、分布式计算与离散数学等高级概念构建技术术语,通过运用富有生产力的梵语帕尼尼语法体系成功解决了这一问题。另一难点是泰卢固语数学排版工具不完善,作者为此开发了名为TeluguTeX的泰卢固语XeLaTeX模板。基于此经验,文章提出面向印地语等超过十亿母语者的印地语族科学写作改善愿景:通过深化梵语技术词汇库建设与推进技术国际化。

原文摘要 · Abstract (English)

I describe my experience writing the first original, modern Computer Science research paper expressed entirely in an Indian language. The paper is in Telugu, a language with approximately 100 million speakers. The paper is in the field of distributed computing and it introduces a technique for proving epistemic logic based lower bounds for multiprocessor algorithms. A key hurdle to writing the paper was developing technical terminology for advanced computer science concepts, including those in algorithms, distributed computing, and discrete mathematics. I overcame this challenge by deriving and coining native language scientific terminology through the powerful, productive, Pāninian grammar of Samskrtam. The typesetting of the paper was an additional challenge, since mathematical typesetting in Telugu is underdeveloped. I overcame this problem by developing a Telugu XeLaTeX template, which I call TeluguTeX. Leveraging this experience of writing an original computer science research paper in an Indian language, I lay out a vision for how to ameliorate the state of scientific writing at all levels in Indic languages -- languages whose native speakers exceed one billion people -- through the further development of the Sanskrit technical lexicon and through technological internationalization.

语言技术科研本土化梵语应用

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。