arXiv:2607.28237cs.AI2026-07

评估AI在伊斯兰经文、圣训等领域的可靠性,发现其适合作为入门辅助工具但不可直接用于宗教裁决。

AI and Authenticity in Islamic Research: A Critical Evaluation of Generative AI Reliability, Hallucination, and Source Fidelity in Quranic, Hadith, and Fiqh Knowledge

论文配图:AI and Authenticity in Islamic Research: A Critical Evaluation of Generative AI Reliability, Hallucination, and Source Fidelity in Quranic, Hadith, and Fiqh Knowledge
图 1 · 摘自论文原文
  • 用50个真实伊斯兰问题测试6大AI系统,覆盖经文、圣训、教法等多个领域
  • 多数AI存在幻觉、引用不全或无法验证,尤其在宗教学派差异问题上表现不一
  • 适合初学者学习辅助,但宗教裁决和学术研究仍需依赖权威文献与专家判断

生成式人工智能正被穆斯林广泛用于宗教指导、古兰经释义、圣训解释、教法裁决及伊斯兰教育。尽管使用日益普遍,但当前缺乏实证证据证明其在高信任宗教场景中能否提供真实、可验证且可信的伊斯兰知识。本研究通过50个现实开放性伊斯兰问题,评估六款主流生成式AI系统的表现,涵盖古兰经诠释、圣训、教法、伦理、牧灵建议及宗教学派敏感议题。数据采集自澳大利亚与英国的参与者,在真实使用情境下收集,并采用混合方法框架分析领域准确性、引文验证、幻觉、教法一致性、不确定性处理、来源溯源及地理差异。研究回答四个核心问题:(1) AI在主要伊斯兰知识领域中的准确性和真实性如何?(2) AI产生幻觉、引用不全或不可验证宗教引用的程度如何?(3) 模型对教法分歧、学派多样性及不确定性的处理是否一致?(4) 当前AI系统是否足够可靠用于宗教指导、伊斯兰教育与学术研究?结果表明,当前生成式AI在初级伊斯兰学习中具有辅助价值,但不应作为宗教裁决或学术研究的权威依据,须对照认证的一手资料与专业学者意见进行验证。本研究是首个针对伊斯兰知识领域中AI可靠性进行的综合性实证评估,为研究人员、教育者、AI开发者及广大穆斯林社区提供实用指引。

原文摘要 · Abstract (English)

Generative Artificial Intelligence (AI) is increasingly used by Muslims for religious guidance, Qur'anic interpretation, Hadith explanation, jurisprudential rulings, and Islamic education. Despite its growing adoption, there is limited empirical evidence on whether current AI systems provide authentic, verifiable, and trustworthy Islamic knowledge suitable for high-trust religious contexts. This study evaluates six leading generative AI systems using fifty realistic open-ended Islamic questions covering Qur'anic interpretation, Hadith, Fiqh, ethics, pastoral advice, and Madhhab-sensitive topics. Responses were collected under real-world conditions from participants in Australia and the United Kingdom and analysed using a mixed-method framework examining domain accuracy, citation verification, hallucinations, jurisprudential consistency, uncertainty handling, source provenance, and geographical variation. The study addresses four research questions: (1) How accurate and authentic are AI-generated responses across major Islamic knowledge domains? (2) To what extent do AI systems produce hallucinations, incomplete citations, or unverifiable religious references? (3) How consistently do models handle jurisprudential disagreement, Madhhab diversity, and uncertainty? (4) Are current AI systems sufficiently reliable for religious guidance, Islamic education, and scholarly research? Overall, current generative AI systems are valuable as assistive tools for introductory Islamic learning but should not be treated as authoritative sources for religious rulings or Islamic research without verification against authenticated primary sources and qualified scholarly expertise. This study provides one of the first comprehensive empirical evaluations of AI reliability within Islamic knowledge, offering practical guidance for researchers, educators, AI developers, and the wider Muslim community.

AI可靠性伊斯兰研究幻觉检测生成式AI

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。