arXiv:2606.01361cs.CV2026-06

用扩散模型让云朵变动物,帮人发现藏在云里的形状

Diamonds in the Sky: Pareidolic Animals in Clouds

论文配图:Diamonds in the Sky: Pareidolic Animals in Clouds
图 1 · 摘自论文原文
  • 用扩散模型将云朵片段转为类动物形态
  • 生成图像可准确预测人们会看到的动物
  • 动态过渡视频提升人们对隐藏动物的识别

人们常在云中看到动物形状,这种现象称为视觉错觉(pareidolia)。现有识别方法通常无法捕捉此类感知。本文提出一种基于AI的方法,可预测人们可能从云中感知到的动物,并帮助未察觉者识别特定错觉动物。该方法利用扩散模型将云朵片段转换为视觉上类似原始云形的动物形象,其成功依赖于目标动物与云形的相似性,且微弱视觉提示即可引导识别。生成图像用于预测潜在感知动物;同时,制作一段从生成图像回溯至原云段的渐变视频,进一步增强人类对错觉动物的感知效果。

原文摘要 · Abstract (English)

People often see animal shapes in clouds, a phenomenon known as pareidolia. We propose an AI-based method that aims to predict which animals people are likely to perceive in clouds, even though state-of-the-art recognition methods typically fail to detect such animals. Additionally, we introduce a method to assist individuals in perceiving specific pareidolic animals, even if they did not recognize them initially. Our approach uses a diffusion model to transform cloud segments into an animal shape that visually resemble the original cloud. This diffusion technique is inspired by the observation that the diffusion process succeeds only when the target animal resembles the shape of the cloud, and that subtle visual hints often suffice to help individuals recognize specific pareidolic animals. A generated image, successfully derived from the diffusion model, is then used to predict the pareidolic animal. Additionally, a short morphing video transitioning from the generated image back to the original cloud segment is employed to further enhance the human's perception of the pareidolic animals.

视觉错觉扩散模型云识别

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。