arXiv:2602.16231cs.CV2026-02

用自然语言精准检索海量视频,自动构建定制化数据集。

DataCube: A Video Retrieval Platform via Natural Language Semantic Profiling

  • 通过语义画像自动生成视频结构化表示
  • 支持神经重排序与深层语义匹配的混合检索
  • 适合需要高效筛选视频数据的研究者与开发者

大规模视频库日益丰富,但将其转化为高质量、任务专用的数据集仍成本高且效率低。我们提出 DataCube,一个智能化平台,实现视频的自动处理、多维语义画像与按需检索。DataCube 构建视频片段的结构化语义表征,支持神经重排序与深度语义匹配的混合检索方式。通过交互式网页界面,用户可从海量视频库中高效构建个性化视频子集,用于训练、分析与评估,并可为私有视频集合建立可搜索系统。系统已公开访问:https://datacube.baai.ac.cn/。演示视频:https://baai-data-cube.ks3-cn-beijing.ksyuncs.com/custom/Adobe%20Express%20-%202%E6%9C%8818%E6%97%A5%20%281%29%281%29%20%281%29.mp4

原文摘要 · Abstract (English)

Large-scale video repositories are increasingly available for modern video understanding and generation tasks. However, transforming raw videos into high-quality, task-specific datasets remains costly and inefficient. We present DataCube, an intelligent platform for automatic video processing, multi-dimensional profiling, and query-driven retrieval. DataCube constructs structured semantic representations of video clips and supports hybrid retrieval with neural re-ranking and deep semantic matching. Through an interactive web interface, users can efficiently construct customized video subsets from massive repositories for training, analysis, and evaluation, and build searchable systems over their own private video collections. The system is publicly accessible at https://datacube.baai.ac.cn/. Demo Video: https://baai-data-cube.ks3-cn-beijing.ksyuncs.com/custom/Adobe%20Express%20-%202%E6%9C%8818%E6%97%A5%20%281%29%281%29%20%281%29.mp4

视频检索语义画像数据构建智能平台

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。