arXiv:2603.13695cs.HCcs.CV2026-03

用AI生成清晰易懂的无障碍图像,降低制作成本。

Steering Generative Models for Accessibility: EasyRead Image Generation

  • 微调Stable Diffusion模型,用LoRA适配器生成统一风格图像。
  • 提出EasyRead评分标准,验证生成图像在清晰度与一致性上的提升。
  • 适合无障碍设计、教育出版等需要简化视觉内容的场景。

EasyRead图示是简洁明了、便于理解的图像,有助于智力障碍者、低读写能力人群或语言障碍者理解信息。传统上,大规模制作此类内容受限于人工设计的成本与专业性。相比之下,自动图像生成可显著降低时间和成本,提升数字与纸质材料的可访问性。然而,基于扩散模型的生成图像常过于复杂且随机种子间风格不一致,难以满足清晰统一的图示需求。为此,本文提出一个统一的生成流程:在多数据集增强样本上,使用LoRA适配器微调Stable Diffusion模型。由于缺乏统一定义,我们引入EasyRead评分以评估图像质量与一致性。实验表明,该方法能有效引导生成具有一致风格和高可读性的图像,证明生成模型可作为规模化、无障碍图示生产的实用工具。

原文摘要 · Abstract (English)

EasyRead pictograms are simple, visually clear images that represent specific concepts and support comprehension for people with intellectual disabilities, low literacy, or language barriers. The large-scale production of EasyRead content has traditionally been constrained by the cost and expertise required to manually design pictograms. In contrast, automatic generation of such images could significantly reduce production time and cost, enabling broader accessibility across digital and printed materials. However, modern diffusion-based image generation models tend to produce outputs that exhibit excessive visual detail and lack stylistic stability across random seeds, limiting their suitability for clear and consistent pictogram generation. This challenge highlights the need for methods specifically tailored to accessibility-oriented visual content. In this work, we present a unified pipeline for generating EasyRead pictograms by fine-tuning a Stable Diffusion model using LoRA adapters on a curated corpus that combines augmented samples from multiple pictogram datasets. Since EasyRead pictograms lack a unified formal definition, we introduce an EasyRead score to benchmark pictogram quality and consistency. Our results demonstrate that diffusion models can be effectively steered toward producing coherent EasyRead-style images, indicating that generative models can serve as practical tools for scalable and accessible pictogram production.

图像生成无障碍设计扩散模型LoRA

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。