梳理160多个目标识别数据集,帮研究者选对基准测试工具。
Object Recognition Datasets and Challenges: A Review
- 系统分析160+公开数据集的特征与适用场景
- 归纳主流评测基准和常用评估指标
- 适合刚入行或需对比模型性能的研究者
目标识别是计算机视觉中的基础任务,支撑着各类图像理解应用。在目标识别研究的各个阶段,都伴随着新数据集的构建与标注,以匹配先进算法的能力。近年来,随着深度网络技术的发展,数据集的规模与质量愈发重要。数据集不仅为竞赛提供公平的评估标准,更推动了目标识别研究的进步。本文对广泛使用的公共数据集进行了详细分析,涵盖超过160个数据集的统计与描述,并综述了主要的目标识别基准与竞赛,以及社区常用的评估指标。所有数据集与挑战信息均可在 github.com/AbtinDjavadifar/ORDC 获取。
原文摘要 · Abstract (English)
Object recognition is among the fundamental tasks in the computer vision applications, paving the path for all other image understanding operations. In every stage of progress in object recognition research, efforts have been made to collect and annotate new datasets to match the capacity of the state-of-the-art algorithms. In recent years, the importance of the size and quality of datasets has been intensified as the utility of the emerging deep network techniques heavily relies on training data. Furthermore, datasets lay a fair benchmarking means for competitions and have proved instrumental to the advancements of object recognition research by providing quantifiable benchmarks for the developed models. Taking a closer look at the characteristics of commonly-used public datasets seems to be an important first step for data-driven and machine learning researchers. In this survey, we provide a detailed analysis of datasets in the highly investigated object recognition areas. More than 160 datasets have been scrutinized through statistics and descriptions. Additionally, we present an overview of the prominent object recognition benchmarks and competitions, along with a description of the metrics widely adopted for evaluation purposes in the computer vision community. All introduced datasets and challenges can be found online at github.com/AbtinDjavadifar/ORDC.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。