arXiv:2509.01072eess.IVcs.AI2025-09被引 2

用物理模型增强眼底图,提升糖尿病视网膜病变诊断准确率

DRetNet: A Novel Deep Learning Framework for Diabetic Retinopathy Diagnosis

  • 结合物理规律动态增强图像,突出微动脉瘤等关键病灶
  • 融合深度特征与人工设计特征,准确率达92.7%、AUC达97.8%
  • 分阶段分类并给出置信度,医生评分4.8/5,适合临床使用

糖尿病视网膜病变是全球失明主因,需早期发现以防止视力丧失。现有自动化检测系统常受限于低质量图像、可解释性差及缺乏领域知识整合。本文提出新型框架,包含三项创新:(1) 基于物理信息神经网络(PINNs)的自适应视网膜图像增强,通过引入物理约束动态提升图像质量,增强微动脉瘤、出血和渗出等关键特征的可见性;(2) 混合特征融合网络(HFFN),融合深度学习嵌入与手工特征,结合学习表示与领域知识,提升泛化能力与精度;(3) 具有不确定性量化的多阶段分类器,将分类过程分解为逻辑阶段,提供可解释预测与置信度评分,增强临床可信度。该框架在测试中达到准确率92.7%、精确率92.5%、召回率92.6%、F1分数92.5%、AUC 97.8%、mAP 0.96、MCC 0.85。眼科医生对预测结果的临床相关性评分高达4.8/5,表明其契合真实诊疗需求。定性分析如Grad-CAM可视化与不确定性热图进一步提升系统可解释性。框架在低质量图像、噪声数据及未见数据集上均表现稳健,具备在资源有限环境推广的潜力。

原文摘要 · Abstract (English)

Diabetic retinopathy (DR) is a leading cause of blindness worldwide, necessitating early detection to prevent vision loss. Current automated DR detection systems often struggle with poor-quality images, lack interpretability, and insufficient integration of domain-specific knowledge. To address these challenges, we introduce a novel framework that integrates three innovative contributions: (1) Adaptive Retinal Image Enhancement Using Physics-Informed Neural Networks (PINNs): this technique dynamically enhances retinal images by incorporating physical constraints, improving the visibility of critical features such as microaneurysms, hemorrhages, and exudates; (2) Hybrid Feature Fusion Network (HFFN): by combining deep learning embeddings with handcrafted features, HFFN leverages both learned representations and domain-specific knowledge to enhance generalization and accuracy; (3) Multi-Stage Classifier with Uncertainty Quantification: this method breaks down the classification process into logical stages, providing interpretable predictions and confidence scores, thereby improving clinical trust. The proposed framework achieves an accuracy of 92.7%, a precision of 92.5%, a recall of 92.6%, an F1-score of 92.5%, an AUC of 97.8%, a mAP of 0.96, and an MCC of 0.85. Ophthalmologists rated the framework's predictions as highly clinically relevant (4.8/5), highlighting its alignment with real-world diagnostic needs. Qualitative analyses, including Grad-CAM visualizations and uncertainty heatmaps, further enhance the interpretability and trustworthiness of the system. The framework demonstrates robust performance across diverse conditions, including low-quality images, noisy data, and unseen datasets. These features make the proposed framework a promising tool for clinical adoption, enabling more accurate and reliable DR detection in resource-limited settings.

眼科医学深度学习可解释性糖尿病视网膜病变

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。