让视障用户通过问答方式探索3D模型,提升可访问性。
SweeperBot: Making 3D Browsing Accessible through View Analysis and Visual Question Answering
- 结合最优视角选择与生成/识别模型,实现视觉问答。
- 10位视障用户实测验证系统可行性,30位明眼人确认描述质量。
- 适合视障群体、3D内容创作者及无障碍设计研究者。
视障用户访问3D模型仍存在困难。尽管部分现有3D查看器允许创作者添加替代文本,但通常对3D模型的描述不够详尽。基于一项形成性研究,本文提出SweeperBot系统,使屏幕阅读器(SR)用户可通过视觉问答探索和比较3D模型。SweeperBot通过结合最优视角选择技术与生成式及识别型基础模型的优势,回答用户提出的视觉问题。10位有屏幕阅读器使用经验的盲人及低视力(BLV)用户专家评审证实了该系统在辅助探索和比较3D模型方面的可行性。其生成描述的质量通过另一项包含30位明眼参与者的调查研究得到验证。
原文摘要 · Abstract (English)
Accessing 3D models remains challenging for Screen Reader (SR) users. While some existing 3D viewers allow creators to provide alternative text, they often lack sufficient detail about the 3D models. Grounded on a formative study, this paper introduces SweeperBot, a system that enables SR users to leverage visual question answering to explore and compare 3D models. SweeperBot answers SR users' visual questions by combining an optimal view selection technique with the strength of generative- and recognition-based foundation models. An expert review with 10 Blind and Low-Vision (BLV) users with SR experience demonstrated the feasibility of using SweeperBot to assist BLV users in exploring and comparing 3D models. The quality of the descriptions generated by SweeperBot was validated by a second survey study with 30 sighted participants.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。