arXiv:2506.15029cs.SDcs.CL2025-06被引 1

用LabVIEW实现高精度OCR语音合成,帮视障者听书

An accurate and revised version of optical character recognition-based speech synthesis using LabVIEW

  • 基于LabVIEW构建OCR语音合成系统
  • 可低成本实现文字转语音,提升可读性
  • 适合视障人士及无障碍教育场景

通过声音获取知识是一种独特能力。视障人士通常仅依赖盲文书籍和非政府组织提供的音频录音,但这些方式存在局限,导致他们难以获取自己想读的书籍。相比文本,声音是视障人士更有效的沟通方式,因其能轻松响应声音信号。本文提出一种准确、可靠、成本低且用户友好的基于光学字符识别(OCR)的语音合成系统,该系统采用实验室虚拟仪器工程工作台(LabVIEW)实现。

原文摘要 · Abstract (English)

Knowledge extraction through sound is a distinctive property. Visually impaired individuals often rely solely on Braille books and audio recordings provided by NGOs. Due to limitations in these approaches, blind individuals often cannot access books of their choice. Speech is a more effective mode of communication than text for blind and visually impaired persons, as they can easily respond to sounds. This paper presents the development of an accurate, reliable, cost-effective, and user-friendly optical character recognition (OCR)-based speech synthesis system. The OCR-based system has been implemented using Laboratory Virtual Instrument Engineering Workbench (LabVIEW).

语音合成OCR无障碍

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。