用手机AI帮视障者识别物体、读取信息并紧急求助
VisionAssist: An Open-Source Smartphone Assistant for AI-Based Visual Accessibility

- 集成实时物体定位、图像语音描述和日程提醒功能
- 支持语音操控,反馈全用语音播报,无需触屏
- 开源项目,便于社区共建无障碍技术
视障人士在完成需视觉判断的日常任务时常遇困难。我们提出 extbf{VisionAssist},一款开源智能手机应用,通过手机摄像头提供AI驱动的视觉辅助,提升独立性。该应用在单个界面集成三项互补功能:一是通过实时摄像头画面定位特定物体;二是对拍摄图像生成语音描述,帮助识别食品标签、文件及日常物品;三是对接手机通讯录与日历,支持紧急呼叫与语音提醒。应用支持语音命令实现免手操作,所有反馈均通过文本转语音输出,确保对视障用户完全可访问。通过将多种辅助服务整合至统一平台,并以开源形式发布,本方案旨在推动社区协作,加速无障碍技术发展。源代码公开于:https://github.com/AOzlemC/LowVisionProject.git
原文摘要 · Abstract (English)
People with low vision often face challenges in performing everyday tasks that require interpreting visual information. We present \textbf{VisionAssist}, an open-source mobile application designed to improve independence by providing AI-powered visual assistance through a smartphone. The application integrates three complementary functionalities within a single interface. First, it enables users to locate specific objects by analyzing the live camera feed. Second, it generates spoken descriptions of captured images, allowing users to identify visual content such as food labels, documents, and everyday objects. Third, it integrates with the smartphone's contacts and calendar to facilitate emergency calls and provide voice-based reminders. The application supports hands-free interaction through voice commands and delivers all feedback using text-to-speech synthesis, making it fully accessible to users with visual impairments. By combining multiple assistive services into a unified platform and releasing the project as open-source software, the proposed solution aims to encourage community contributions and accelerate the development of accessible technologies. The source code is publicly available at: https://github.com/AOzlemC/LowVisionProject.git
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。