arXiv:2606.02481cs.CV2026-06

高分辨率野外场景数据集,支持真实视觉研究

Places in the Wild: A Large, High-Resolution RAW Photograph Dataset for Ecologically Valid Vision Research

  • 在810个真实地点采集67,574张4500万像素RAW照片
  • 每处拍摄72张水平+12张俯仰图像,实现360度密集采样
  • 含完整元数据和图像质量指标,适合真实场景研究

大型图像数据集推动了认知神经科学与计算机视觉的发展。然而,多数数据集为低分辨率、互联网获取的JPEG图片,拍摄条件未知且空间上下文有限。本文提出Places in the Wild,包含67,574张高分辨率照片,实地采集于810个物理位置,覆盖260类基本场景(包括室内、城市与自然环境)。每个地点使用安装在全景三脚架上的4500万像素佳能EOS R5相机,以5度水平间隔拍摄72张图像,并补充12张不同仰角图像,实现密集的360度视角采样。所有图像同步记录为14位RAW(CR3)文件与压缩JPEG,保留传感器级细节,可用于亮度、对比度、颜色等图像统计分析。数据集附带完整的EXIF元数据及一系列图像质量指标,支持人类与模型的视点依赖识别研究、场景理解系统在真实条件下的训练与评估、自然场景统计特征刻画,以及需要近全视野视觉呈现的实验。

原文摘要 · Abstract (English)

Large image datasets have accelerated progress in cognitive neuroscience and computer vision. However, most datasets are low-resolution, internet-sourced JPEGs with unknown capture conditions and limited spatial context. Places in the Wild is a dataset of 67,574 high-resolution photographs collected in situ across 810 physical locations spanning 260 basic-level scene categories, including indoor, urban, and natural environments. At each location, a 45-megapixel Canon EOS R5 mounted on a panoramic tripod captured 72 images at 5-degree horizontal intervals plus 12 images at varying elevations, yielding dense 360-degree viewpoint sampling. All images were recorded simultaneously as 14-bit RAW (CR3) files and compressed JPEGs, preserving sensor-level detail for analyses of luminance, contrast, color, and other image statistics. The dataset is accompanied by complete EXIF metadata and a suite of image-quality metrics. Places in the Wild supports research on viewpoint-dependent recognition in humans and models, training and evaluation of scene-understanding systems under realistic conditions, characterization of natural scene statistics, and experiments requiring near-full-field visual displays.

场景理解高分辨率野外数据图像统计

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。